lancedb

mirror of https://github.com/lancedb/lancedb.git synced 2025-12-26 14:49:57 +00:00

Author	SHA1	Message	Date
Will Jones	5f6d13e958	ci: lint and enforce linting (#829 ) @eddyxu added instructions for linting here: `7af213801a/python/README.md (L45-L50)` However, we had a lot of failures and weren't checking this in CI. This PR fixes all lints and adds a check to CI to keep us in compliance with the lints.	2024-04-05 16:27:31 -07:00
Chang She	72b39432e8	feat(python): add exist_ok option to create table (#813 ) This mimics CREATE TABLE IF NOT EXISTS behavior. We add `db.create_table(..., exist_ok=True)` parameter. By default it is set to False, so trying to create a table with the same name will raise an exception. If set to True, then it only opens the table if it already exists. If you pass in a schema, it will be checked against the existing table to make sure you get what you want. If you pass in data, it will NOT be added to the existing table.	2024-04-05 16:26:35 -07:00
Rok Mihevc	78ab9068a8	feat(python): expose index cache size (#655 ) This is to enable https://github.com/lancedb/lancedb/issues/641. Should be merged after https://github.com/lancedb/lance/pull/1587 is released.	2024-04-05 16:23:49 -07:00
Lei Xu	86efd36689	chore: improve create_table API consistency between local and remote SDK (#627 )	2024-04-05 16:23:47 -07:00
Bert	cd7a4dd251	fix!: sort table names (#619 ) https://github.com/lancedb/lance/issues/1385	2024-04-05 16:22:59 -07:00
Weston Pace	301e08f30e	feat: allow prefiltering with index (#610 ) Support for prefiltering with an index was added in lance version 0.8.7. We can remove the lancedb check that prevents this. Closes #261	2024-04-05 16:22:59 -07:00
Chang She	8469d010f8	feat: add to_list and to_pandas api's (#556 ) Add `to_list` to return query results as list of python dict (so we're not too pandas-centric). Closes #555 Add `to_pandas` API and add deprecation warning on `to_df`. Closes #545 Co-authored-by: Chang She <chang@lancedb.com>	2024-04-05 16:22:59 -07:00
Chang She	c21f9cdda0	ci: fix docs build (#496 ) python/python.md contains typos in the class references --------- Co-authored-by: Chang She <chang@lancedb.com>	2023-09-18 13:07:21 -07:00
Lei Xu	b315ea3978	[Python] Pydantic vector field with default value (#474 ) Rename `lance.pydantic.vector` to `Vector` and deprecate `vector(dim)`	2023-09-08 22:35:31 -07:00
Ayush Chaurasia	aa7806cf0d	[Python]Fix record_batch_generator (#483 ) Should fix - https://github.com/lancedb/lancedb/issues/482	2023-09-08 21:18:50 +05:30
Chang She	9a9a73a65d	[python] Use pydantic for embedding function persistence (#467 ) 1. Support persistent embedding function so users can just search using query string 2. Add fixed size list conversion for multiple vector columns 3. Add support for empty query (just apply select/where/limit). 4. Refactor and simplify some of the data prep code --------- Co-authored-by: Chang She <chang@lancedb.com> Co-authored-by: Weston Pace <weston.pace@gmail.com>	2023-09-05 21:30:45 -07:00
Ayush Chaurasia	0b9924b432	Make creating (and adding to) tables via Iterators more flexible & intuitive (#430 ) It improves the UX as iterators can be of any type supported by the table (plus recordbatch) & there is no separate requirement. Also expands the test cases for pydantic & arrow schema. If this is looks good I'll update the docs. Example usage: ``` class Content(LanceModel): vector: vector(2) item: str price: float def make_batches(): for _ in range(5): yield from [ # pandas pd.DataFrame({ "vector": [[3.1, 4.1], [1, 1]], "item": ["foo", "bar"], "price": [10.0, 20.0], }), # pylist [ {"vector": [3.1, 4.1], "item": "foo", "price": 10.0}, {"vector": [5.9, 26.5], "item": "bar", "price": 20.0}, ], # recordbatch pa.RecordBatch.from_arrays( [ pa.array([[3.1, 4.1], [5.9, 26.5]], pa.list_(pa.float32(), 2)), pa.array(["foo", "bar"]), pa.array([10.0, 20.0]), ], ["vector", "item", "price"], ), # pydantic list [ Content(vector=[3.1, 4.1], item="foo", price=10.0), Content(vector=[5.9, 26.5], item="bar", price=20.0), ]] db = lancedb.connect("db") tbl = db.create_table("tabley", make_batches(), schema=Content, mode="overwrite") tbl.add(make_batches()) ``` Same should with arrow schema. --------- Co-authored-by: Weston Pace <weston.pace@gmail.com>	2023-08-18 09:56:30 +05:30
Ashis Kumar Naik	902a402951	implementation of drop_database (#418 ) #416 Fixed. added drop_database() method . This deletes all the tables from the database with a single command. --------- Signed-off-by: Ashis Kumar Naik <ashishami2002@gmail.com>	2023-08-11 20:59:56 -07:00
Chang She	a54d1e5618	Automatically convert pydantic model (#400 ) Saves users from having to explicitly call `LanceModel.to_arrow_schema()` when creating an empty table. See new docs for full details. --------- Co-authored-by: Chang She <chang@lancedb.com>	2023-08-06 14:50:03 -07:00
Ayush Chaurasia	bbfadfe58d	[python] Allow adding via iterators (#391 ) Makes the following work so all the formats accepted by `create_table()` are also accepted by `add()` ``` import lancedb import pyarrow as pa db = lancedb.connect("/tmp") def make_batches(): for i in range(5): yield pa.RecordBatch.from_arrays( [ pa.array([[3.1, 4.1], [5.9, 26.5]]), pa.array(["foo", "bar"]), pa.array([10.0, 20.0]), ], ["vector", "item", "price"], ) schema = pa.schema([ pa.field("vector", pa.list_(pa.float32())), pa.field("item", pa.utf8()), pa.field("price", pa.float32()), ]) tbl = db.create_table("table4", make_batches(), schema=schema) tbl.add(make_batches()) ```	2023-08-04 12:49:44 -07:00
Chang She	2d25c263e9	Implement drop table if exists (#383 )	2023-07-31 10:25:09 +02:00
Lei Xu	088e745e1d	[Python] Create table with Iterator[RecordBatch] and add docs (#316 )	2023-07-16 21:45:55 -07:00
Chang She	2fdcb307eb	[python] Fix a few minor bugs (#304 )	2023-07-15 03:47:42 +08:00
Leon Yee	eb5bcda337	Error implementations (#232 ) Solves #216 by adding a check on table open for existence of the `.lance` file. Does not check for it for remote connections.	2023-06-27 16:48:31 -07:00
Lei Xu	4bc676e26a	[Python] Support replace during create_index (#233 ) Closes #214	2023-06-27 16:02:07 -07:00
Tevin Wang	9b83ce3d2a	add black to python CI (#178 ) Closes #48	2023-06-12 11:22:34 -07:00
Lei Xu	9965b4564d	[Python] Support drop table (#123 ) Closes #86	2023-06-01 15:58:45 -07:00
Chang She	59014a01e0	bump version for v0.1.2	2023-05-05 11:27:09 -07:00
Chang She	d7c5793803	Add mode to overwrite table if already exists	2023-04-19 16:22:11 -07:00
Lei Xu	c38d80cab2	remove print	2023-04-19 14:26:07 -07:00
Lei Xu	b3fdabdf45	use python and arrow	2023-04-19 14:15:18 -07:00
Chang She	5ef5141812	black	2023-03-22 18:29:07 -07:00
Chang She	690141d357	add unit tests	2023-03-21 22:29:19 -07:00
Chang She	b10301f5d6	initial python impl	2023-03-18 10:43:26 -07:00

29 Commits