A supercharged SQLite library for Python

Last update: Dec 30, 2022

Related tags

Overview

SuperSQLite: a supercharged SQLite library for Python

A feature-packed Python package and for utilizing SQLite in Python by Plasticity. It is intended to be a drop-in replacement to Python's built-in SQLite API, but without any limitations. It offers unique features like remote streaming over HTTP and bundling of extensions like JSON, R-Trees (geospatial indexing), and Full Text Search. SuperSQLite is also packaged with pre-compiled native binaries for SQLite and all of its extensions for nearly every platform as to avoid any C/C++ compiler errors during install.

Installation
Motivation
Using the Library
- Connecting
- Querying
Remote Streaming over HTTP
Other Documentation
Other Programming Languages
Contributing
Roadmap
Other Notable Projects
LICENSE and Attribution

Installation

You can install this package with pip:

pip install supersqlite # Python 2.7
pip3 install supersqlite # Python 3

Motivation

SQLite, is a fast, popular embedded database, used by large enterprises. It is the most widely-deployed database and has billions of deployments. It has a built-in binding in Python.

The Python bindings, however, often are compiled against an out-of-date copy of SQLite or may be compiled with limitations set to low levels. Moreover, it is difficult to load extremely useful extensions like JSON1 that adds JSON functionality to SQLite or FTS5 that adds full-text search functionality to SQLite since they must be compiled with a C/C++ compiler on each platform before being loaded.

SuperSQLite aims to solve these problems by packaging a newer version of SQLite natively pre-compiled for every platform along with natively pre-compiled SQLite extensions. SuperSQLite also adds useful unique new features like remote streaming over HTTP to read from a centralized SQLite database.

Moreover, by default, SQLite does not enable some optimizations that can result in speedups. SuperSQLite compiles SQLite with various optimizations and allows you to select your workload at runtime to further automatically configure the connection to be optimized for your workload.

When to use SuperSQLite?

SQLite is extremely reliable and durable for large amounts of data (up to 140TB). It is considered one of the most well-engineered and well-tested software solutions today, with 711x more test code than implementation code.

SQLite is faster than nearly every other database at read-heavy use cases (especially compared to databases that may use a client-server model with network latency like MySQL, PostgreSQL, MongoDB, DynamoDB, etc.). You can also instantiate SQLite completely in-memory to remove disk latency, if your data will fit within RAM. For key/value use cases, you can get comparable or better read/write performance to key/value databases like LevelDB with the LSM1 extension.

When you have a write-heavy workload with multiple servers that need to write concurrently to a shared database (backend to a website), you would probably want to choose something that has a client-server model instead like PostgreSQL, although SQLite can handle processing write requests fast enough that it is sufficient for most concurrent write loads. In fact, Expensify uses SQLite for their entire backend. If you need the database to be automatically replicated or automatically sharded across machines or other distributed features, you probably want to use something else.

See Appropriate Uses For SQLite for more information and Well-Known Users of SQLite for example use cases.

Using the Library

Instead of 'import sqlite3', use:

from supersqlite import sqlite3

This retains compatibility with the sqlite3 package, while adding the various enhancements.

Connecting

Given the above import, connect to a sqlite database file using:

conn = sqlite3.connect('foo.db')

Querying

Remote Streaming over HTTP

Workload Optimizations

Extensions

JSON1

FTS3, FTS4, FTS5

LSM1

R*Tree

Other

Custom

Export SQLite Resources

Optimizations

Other Programming Languages

Currently, this library only supports Python. There are no plans to port it to any other languages, but since SQLite has a native C implementation and has bindings in most languages, you can use the export functions to load SuperSQLite's SQLite extensions in the SQLite bindings of other programming languages or link SuperSQLite's version of SQLite to a native binary.

Contributing

The main repository for this project can be found on GitLab. The GitHub repository is only a mirror. Pull requests for more tests, better error-checking, bug fixes, performance improvements, or documentation or adding additional utilties / functionalities are welcome on GitLab.

You can contact us at [email protected].

Roadmap

Out of the box, "fast-write" configuration option that makes the connection optimized for fast-writing.
Out of the box, "fast-read" configuration option that makes the conenction optimized for fast-reading.
Optimize streaming cache behavior

Other Notable Projects

pysqlite - The built-in sqlite3 module in Python.
apsw - Powers the main API of SuperSQLite, aims to port all of SQLite's API functionality (like VFSes) to Python, not just the query APIs.
Magnitude - Another project by Plasticity that uses SuperSQLite's unique features for machine learning embedding models.

LICENSE and Attribution

This repository is licensed under the license found here.

The SQLite "feather" icon is taken from the SQLite project which is released as public domain.

This project is not affiliated with the official SQLite project.

Comments

Missing __enter__() with cursor

The syntax : with conn.cursor() as cursor: ... work with standard Python SQLite and others database drivers (postgres, ...) But the method __enter__() is not defined in supersqlite driver.

opened by pprados 0
AttributeError: 'sqlite3.Connection' object has no attribute 'enable_load_extension'

Is supersqlite enable loading sqlite extensions? What can cause this error:

from supersqlite import sqlite3 as sqlite33 conn33=sqlite33.connect("mydbfile.db") conn33.enable_load_extension(True) Traceback (most recent call last): File "", line 1, in AttributeError: 'sqlite3.Connection' object has no attribute 'enable_load_extension' thank you.

opened by dbricker-intel 0
Publish supersqlite on conda-forge

It would be very convenient to have this library available on conda-forge so it could be installed with the Conda package manager, which is ideal for packages with binary dependencies. Is that a possibility?

opened by JWCook 0

No module named 'supersqlite.third_party.internal.apsw'

Hello, I try to install supersqlite from pip and conda. I can use

docker run -it ubuntu
# then
apt-get update ; \
apt-get install -y python-apsw python3 python3-pip ; \
pip3 install supersqlite ; python3 -c "import supersqlite"

It's correct.

Now, I try with conda

docker run -it conda/miniconda3
# Then
conda update conda -y ; \
conda init bash ;
exec bash
# and
conda install -c conda-forge apsw -y ; \
pip3 install supersqlite ; \
python3 -c "import supersqlite"

I receive and error: ModuleNotFoundError: No module named 'supersqlite.third_party.internal.apsw'

Collecting supersqlite
  Downloading supersqlite-0.0.78.tar.gz (25.8 MB)
     |████████████████████████████████| 25.8 MB 3.9 MB/s 
Building wheels for collected packages: supersqlite
  Building wheel for supersqlite (setup.py) ... done
  Created wheel for supersqlite: filename=supersqlite-0.0.78-cp39-cp39-linux_x86_64.whl size=71094064 sha256=64d23d0848fba148a4506b0905135d23df4d7099660bbba683584b71623d38a2
  Stored in directory: /root/.cache/pip/wheels/c2/83/40/cffebda33928fae730f81985e5d75078d257db3586bd419905
Successfully built supersqlite
Installing collected packages: supersqlite
Successfully installed supersqlite-0.0.78
Traceback (most recent call last):
  File "<string>", line 1, in <module>
  File "/usr/local/lib/python3.9/site-packages/supersqlite/__init__.py", line 47, in <module>
    import supersqlite.third_party.internal.apsw as apsw
ModuleNotFoundError: No module named 'supersqlite.third_party.internal.apsw'

You can reproduce this bug with this docker file

# SuperSqliteDockerfile file
FROM continuumio/anaconda3

RUN conda create --name testSupersqlite python=3.7 ; \
    conda activate testSupersqlite ; \
    pip3 install supersqlite

ENTRYPOINT python -c "import supersqlite"

and

docker run --rm -it $(docker build -q -f SuperSqliteDockerfile .)

How can I resolve this problem?

Thanks

opened by pprados 1

Fix typo that causes install to break

pip install of supersqlite is broken due to a typo in a requirement name: lsb-db should actually be lsm-db.

This also resolves https://github.com/plasticityai/supersqlite/issues/5

opened by DDevine 2
[BUG] Outdated/typo in requirements.txt
Description

It seems current requirements.txt is either outdated or contains a typo in https://github.com/plasticityai/supersqlite/blob/d74da749c6fa5df021df3968b854b9a59f829e17/requirements.txt#L1 When trying to pip install lsb-db==0.6.4 I get following error

Could not find a version that satisfies the requirement lsb-db==0.6.4 (from versions: none)

I guess this could be a typo for lsm-db?

Expected behavior

Installing this package via pip won't fail

System

Ubuntu 18.04LTS / Python 3.6 / pip 19.1.1

cc @AjayP13
opened by johnygomez 0

Releases(0.0.78)

0.0.78(Nov 22, 2018)

null
Source code(tar.gz)
Source code(zip)
0.0.77(Nov 19, 2018)

null
Source code(tar.gz)
Source code(zip)
0.0.76(Nov 19, 2018)

null
Source code(tar.gz)
Source code(zip)
0.0.75(Nov 19, 2018)

null
Source code(tar.gz)
Source code(zip)
0.0.74(Nov 18, 2018)

null
Source code(tar.gz)
Source code(zip)
0.0.73(Nov 18, 2018)

null
Source code(tar.gz)
Source code(zip)
0.0.72(Nov 16, 2018)

null
Source code(tar.gz)
Source code(zip)
0.0.71(Nov 16, 2018)

null
Source code(tar.gz)
Source code(zip)
0.0.70(Nov 16, 2018)

null
Source code(tar.gz)
Source code(zip)
0.0.69(Nov 16, 2018)

null
Source code(tar.gz)
Source code(zip)
0.0.68(Nov 16, 2018)

null
Source code(tar.gz)
Source code(zip)
0.0.67(Nov 16, 2018)

null
Source code(tar.gz)
Source code(zip)
0.0.66(Nov 16, 2018)

null
Source code(tar.gz)
Source code(zip)
0.0.65(Nov 16, 2018)

null
Source code(tar.gz)
Source code(zip)
0.0.64(Nov 16, 2018)

null
Source code(tar.gz)
Source code(zip)
0.0.63(Nov 16, 2018)

null
Source code(tar.gz)
Source code(zip)
0.0.61(Nov 16, 2018)

null
Source code(tar.gz)
Source code(zip)
0.0.60(Nov 15, 2018)

null
Source code(tar.gz)
Source code(zip)
0.0.59(Nov 15, 2018)

null
Source code(tar.gz)
Source code(zip)
0.0.58(Nov 15, 2018)

null
Source code(tar.gz)
Source code(zip)
0.0.56(Nov 15, 2018)

null
Source code(tar.gz)
Source code(zip)
0.0.55(Nov 15, 2018)

null
Source code(tar.gz)
Source code(zip)
0.0.54(Nov 15, 2018)

null
Source code(tar.gz)
Source code(zip)
0.0.52(Nov 15, 2018)

null
Source code(tar.gz)
Source code(zip)
0.0.51(Nov 15, 2018)

null
Source code(tar.gz)
Source code(zip)
0.0.50(Nov 15, 2018)

null
Source code(tar.gz)
Source code(zip)
0.0.49(Nov 15, 2018)

null
Source code(tar.gz)
Source code(zip)
0.0.47(Nov 15, 2018)

null
Source code(tar.gz)
Source code(zip)
0.0.46(Nov 15, 2018)

null
Source code(tar.gz)
Source code(zip)
0.0.45(Nov 15, 2018)

null
Source code(tar.gz)
Source code(zip)

Owner

Plasticity

The official GitHub account of Plasticity

GitHub Repository

Redis Python Client

redis-py The Python interface to the Redis key-value store. Python 2 Compatibility Note redis-py 3.5.x will be the last version of redis-py that suppo

11k Dec 29, 2022

Familiar asyncio ORM for python, built with relations in mind

Tortoise ORM Introduction Tortoise ORM is an easy-to-use asyncio ORM (Object Relational Mapper) inspired by Django. Tortoise ORM was build with relati

3.3k Dec 31, 2022

Dlsite-doujin-renamer - Dlsite doujin renamer tool with python

dlsite-doujin-renamer Features 支持深度查找带有 RJ 号的文件夹支持手动选择文件夹或拖拽文件夹到软件窗口支持在 config

111 Jan 02, 2023

This is a repository for a task assigned to me by Bilateral solutions!

Processing-Files-using-MySQL This is a repository for a task assigned to me by Bilateral solutions! Task: Make Folders named Processing,queue and proc

1 Nov 07, 2022

This repository is for active development of the Azure SDK for Python.

Azure SDK for Python This repository is for active development of the Azure SDK for Python. For consumers of the SDK we recommend visiting our public

3.4k Jan 02, 2023

A pythonic interface to Amazon's DynamoDB

PynamoDB A Pythonic interface for Amazon's DynamoDB. DynamoDB is a great NoSQL service provided by Amazon, but the API is verbose. PynamoDB presents y

2.1k Dec 30, 2022

A SQL linter and auto-formatter for Humans

The SQL Linter for Humans SQLFluff is a dialect-flexible and configurable SQL linter. Designed with ELT applications in mind, SQLFluff also works with

5.5k Jan 08, 2023

#crypto #cipher #encode #decode #hash

🌹 CYPHER TOOLS 🌹 Written by TMRSWRR Version 1.0.0 All in one tools for CRYPTOLOGY. Instagram: Capture the Root 🖼️ Screenshots 🖼️ 📹 How to use 📹

50 Dec 23, 2022

PostgreSQL database access simplified

Queries: PostgreSQL Simplified Queries is a BSD licensed opinionated wrapper of the psycopg2 library for interacting with PostgreSQL. The popular psyc

251 Oct 25, 2022

A simple python package that perform SQL Server Source Control and Auto Deployment.

deploydb Deploy your database objects automatically when the git branch is updated. Production-ready! ⚙️ Easy-to-use 🔨 Customizable 🔧 Installation I

10 Dec 07, 2022

Python cluster client for the official redis cluster. Redis 3.0+.

redis-py-cluster This client provides a client for redis cluster that was added in redis 3.0. This project is a port of redis-rb-cluster by antirez, w

1.1k Jan 05, 2023

Kafka Connect JDBC Docker Image.

kafka-connect-jdbc This is a dockerized version of the Confluent JDBC database connector. Usage This image is running the connect-standalone command w

1 Jan 05, 2022

Apache Libcloud is a Python library which hides differences between different cloud provider APIs and allows you to manage different cloud resources through a unified and easy to use API

Apache Libcloud - a unified interface for the cloud Apache Libcloud is a Python library which hides differences between different cloud provider APIs

1.9k Dec 25, 2022

Pandas on AWS - Easy integration with Athena, Glue, Redshift, Timestream, QuickSight, Chime, CloudWatchLogs, DynamoDB, EMR, SecretManager, PostgreSQL, MySQL, SQLServer and S3 (Parquet, CSV, JSON and EXCEL).

AWS Data Wrangler Pandas on AWS Easy integration with Athena, Glue, Redshift, Timestream, QuickSight, Chime, CloudWatchLogs, DynamoDB, EMR, SecretMana

3.3k Dec 31, 2022

A supercharged SQLite library for Python

Related tags

Overview

SuperSQLite: a supercharged SQLite library for Python

Table of Contents

Installation

Motivation

When to use SuperSQLite?

Using the Library

Connecting

Querying

Remote Streaming over HTTP

Workload Optimizations

Extensions

JSON1

FTS3, FTS4, FTS5

LSM1

R*Tree

Other

Custom

Export SQLite Resources

Optimizations

Other Documentation

Other Programming Languages

Contributing

Roadmap

Other Notable Projects

LICENSE and Attribution

Comments

Missing __enter__() with cursor

AttributeError: 'sqlite3.Connection' object has no attribute 'enable_load_extension'

Publish supersqlite on conda-forge

No module named 'supersqlite.third_party.internal.apsw'

Fix typo that causes install to break

[BUG] Outdated/typo in requirements.txt

Description

Expected behavior

System

Releases(0.0.78)

0.0.78(Nov 22, 2018)

0.0.77(Nov 19, 2018)

0.0.76(Nov 19, 2018)

0.0.75(Nov 19, 2018)

0.0.74(Nov 18, 2018)

0.0.73(Nov 18, 2018)

0.0.72(Nov 16, 2018)

0.0.71(Nov 16, 2018)

0.0.70(Nov 16, 2018)

0.0.69(Nov 16, 2018)

0.0.68(Nov 16, 2018)

0.0.67(Nov 16, 2018)

0.0.66(Nov 16, 2018)

0.0.65(Nov 16, 2018)

0.0.64(Nov 16, 2018)

0.0.63(Nov 16, 2018)

0.0.61(Nov 16, 2018)

0.0.60(Nov 15, 2018)

0.0.59(Nov 15, 2018)

0.0.58(Nov 15, 2018)

0.0.56(Nov 15, 2018)

0.0.55(Nov 15, 2018)

0.0.54(Nov 15, 2018)

0.0.52(Nov 15, 2018)

0.0.51(Nov 15, 2018)

0.0.50(Nov 15, 2018)

0.0.49(Nov 15, 2018)

0.0.47(Nov 15, 2018)

0.0.46(Nov 15, 2018)

0.0.45(Nov 15, 2018)