LEANN

Author	SHA1	Message	Date
Andy Lee	42690cb74e	docs: remove ollama embedding extra instructions	2025-08-09 16:46:47 -07:00
Andy Lee	1071479c05	docs: Add clear documentation for Ollama embedding usage	2025-08-08 18:09:06 -07:00
yichuan520030910320	67fef60466	[Readme]More about claude code	2025-08-08 16:05:35 -07:00
yichuan520030910320	0b2b799d5a	[README]fix instructions in cli	2025-08-08 01:04:13 -07:00
yichuan520030910320	0f790fbbd9	docs: polish README and add optimized MCP integration image - Improve grammar and sentence structure in MCP section - Add proper markdown image formatting with relative paths - Optimize mcp_leann.png size (1.3MB -> 224KB) - Update data description to be more specific about Chinese content	2025-08-08 00:58:36 -07:00
Andy Lee	5567302316	feat: promote Claude Code integration as primary RAG feature	2025-08-07 23:19:19 -07:00
Andy Lee	8b22d2b5d3	Merge pull request #19 from yichuan-w/feature/claude-code-research Feature/claude code research	2025-08-05 23:02:34 -07:00
Andy Lee	4cb544ee38	docs: Update co-contributors with GitHub usernames (#18 ) * docs: Update co-contributors with GitHub usernames * docs: Use GitHub links for co-contributors and improve order * docs: Change to Contributors and use personal homepage * docs: Specify core contributors and welcome new contributors	2025-08-05 17:43:59 -07:00
yichuan520030910320	f94ce63d51	add gpt oss! serve your RAG using ollama	2025-08-05 16:49:52 -07:00
Andy Lee	0d448c4a41	docs: config guidance (#17 ) * docs: config guidance * feat: add comprehensive configuration guide and update README - Create docs/configuration-guide.md with detailed guidance on: - Embedding model selection (small/medium/large) - Index selection (HNSW vs DiskANN) - LLM engine and model comparison - Parameter tuning (build/search complexity, top-k) - Performance optimization tips - Deep dive into LEANN's recomputation feature - Update README.md to link to the configuration guide - Include latest 2025 model recommendations (Qwen3, DeepSeek-R1, O3-mini) * chore: move evaluation data .gitattributes to correct location * docs: Weaken DiskANN emphasis in README - Change backend description to emphasize HNSW as default - DiskANN positioned as optional for billion-scale datasets - Simplify evaluation commands to be more generic * docs: Adjust DiskANN positioning in features and roadmap - features.md: Put HNSW/FAISS first as default, DiskANN as optional - roadmap.md: Reorder to show HNSW integration before DiskANN - Consistent with positioning DiskANN as advanced option for large-scale use * docs: Improve configuration guide based on feedback - List specific files in default data/ directory (2 AI papers, literature, tech report) - Update examples to use English and better RAG-suitable queries - Change full dataset reference to use --max-items -1 - Adjust small model guidance about upgrading to larger models when time allows - Update top-k defaults to reflect actual default of 20 - Ensure consistent use of full model name Qwen/Qwen3-Embedding-0.6B - Reorder optimization steps, move MLX to third position - Remove incorrect chunk size tuning guidance - Change README from 'Having trouble' to 'Need best practices' * docs: Address all configuration guide feedback - Fix grammar: 'If time is not a constraint' instead of 'time expense is not large' - Highlight Qwen3-Embedding-0.6B performance (nearly OpenAI API level) - Add OpenAI quick start section with configuration example - Fold Cloud vs Local trade-offs into collapsible section - Update HNSW as 'default and recommended for extreme low storage' - Add DiskANN beta warning and explain PQ+rerank architecture - Expand Ollama models: add qwen3:0.6b, 4b, 7b variants - Note OpenAI as current default but recommend Ollama switch - Add 'need to install extra software' warning for Ollama - Remove incorrect latency numbers from search-complexity recommendations * docs: add a link	2025-08-04 22:50:32 -07:00
Andy Lee	8899734952	refactor: Unify examples interface with BaseRAGExample (#12 ) * refactor: Unify examples interface with BaseRAGExample - Create BaseRAGExample base class for all RAG examples - Refactor 4 examples to use unified interface: - document_rag.py (replaces main_cli_example.py) - email_rag.py (replaces mail_reader_leann.py) - browser_rag.py (replaces google_history_reader_leann.py) - wechat_rag.py (replaces wechat_history_reader_leann.py) - Maintain 100% parameter compatibility with original files - Add interactive mode support for all examples - Unify parameter names (--max-items replaces --max-emails/--max-entries) - Update README.md with new examples usage - Add PARAMETER_CONSISTENCY.md documenting all parameter mappings - Keep main_cli_example.py for backward compatibility with migration notice All default values, LeannBuilder parameters, and chunking settings remain identical to ensure full compatibility with existing indexes. * fix: Update CI tests for new unified examples interface - Rename test_main_cli.py to test_document_rag.py - Update all references from main_cli_example.py to document_rag.py - Update tests/README.md documentation The tests now properly test the new unified interface while maintaining the same test coverage and functionality. * fix: Fix pre-commit issues and update tests - Fix import sorting and unused imports - Update type annotations to use built-in types (list, dict) instead of typing.List/Dict - Fix trailing whitespace and end-of-file issues - Fix Chinese fullwidth comma to regular comma - Update test_main_cli.py to test_document_rag.py - Add backward compatibility test for main_cli_example.py - Pass all pre-commit hooks (ruff, ruff-format, etc.) * refactor: Remove old example scripts and migration references - Delete old example scripts (mail_reader_leann.py, google_history_reader_leann.py, etc.) - Remove migration hints and backward compatibility - Update tests to use new unified examples directly - Clean up all references to old script names - Users now only see the new unified interface * fix: Restore embedding-mode parameter to all examples - All examples now have --embedding-mode parameter (unified interface benefit) - Default is 'sentence-transformers' (consistent with original behavior) - Users can now use OpenAI or MLX embeddings with any data source - Maintains functional equivalence with original scripts * docs: Improve parameter categorization in README - Clearly separate core (shared) vs specific parameters - Move LLM and embedding examples to 'Example Commands' section - Add descriptive comments for all specific parameters - Keep only truly data-source-specific parameters in specific sections * docs: Make example commands more representative - Add default values to parameter descriptions - Replace generic examples with real-world use cases - Focus on data-source-specific features in examples - Remove redundant demonstrations of common parameters * docs: Reorganize parameter documentation structure - Move common parameters to a dedicated section before all examples - Rename sections to 'X-Specific Arguments' for clarity - Remove duplicate common parameters from individual examples - Better information architecture for users * docs: polish applications * docs: Add CLI installation instructions - Add two installation options: venv and global uv tool - Clearly explain when to use each option - Make CLI more accessible for daily use * docs: Clarify CLI global installation process - Explain the transition from venv to global installation - Add upgrade command for global installation - Make it clear that global install allows usage without venv activation * docs: Add collapsible section for CLI installation - Wrap CLI installation instructions in details/summary tags - Keep consistent with other collapsible sections in README - Improve document readability and navigation * style: format * docs: Fix collapsible sections - Make Common Parameters collapsible (as it's lengthy reference material) - Keep CLI Installation visible (important for users to see immediately) - Better information hierarchy * docs: Add introduction for Common Parameters section - Add 'Flexible Configuration' heading with descriptive sentence - Create parallel structure with 'Generation Model Setup' section - Improve document flow and readability * docs: nit * fix: Fix issues in unified examples - Add smart path detection for data directory - Fix add_texts -> add_text method call - Handle both running from project root and examples directory * fix: Fix async/await and add_text issues in unified examples - Remove incorrect await from chat.ask() calls (not async) - Fix add_texts -> add_text method calls - Verify search-complexity correctly maps to efSearch parameter - All examples now run successfully * feat: Address review comments - Add complexity parameter to LeannChat initialization (default: search_complexity) - Fix chunk-size default in README documentation (256, not 2048) - Add more index building parameters as CLI arguments: - --backend-name (hnsw/diskann) - --graph-degree (default: 32) - --build-complexity (default: 64) - --no-compact (disable compact storage) - --no-recompute (disable embedding recomputation) - Update README to document all new parameters * feat: Add chunk-size parameters and improve file type filtering - Add --chunk-size and --chunk-overlap parameters to all RAG examples - Preserve original default values for each data source: - Document: 256/128 (optimized for general documents) - Email: 256/25 (smaller overlap for email threads) - Browser: 256/128 (standard for web content) - WeChat: 192/64 (smaller chunks for chat messages) - Make --file-types optional filter instead of restriction in document_rag - Update README to clarify interactive mode and parameter usage - Fix LLM default model documentation (gpt-4o, not gpt-4o-mini) * feat: Update documentation based on review feedback - Add MLX embedding example to README - Clarify examples/data content description (two papers, Pride and Prejudice, Chinese README) - Move chunk parameters to common parameters section - Remove duplicate chunk parameters from document-specific section * docs: Emphasize diverse data sources in examples/data description * fix: update default embedding models for better performance - Change WeChat, Browser, and Email RAG examples to use all-MiniLM-L6-v2 - Previous Qwen/Qwen3-Embedding-0.6B was too slow for these use cases - all-MiniLM-L6-v2 is a fast 384-dim model, ideal for large-scale personal data * add response highlight * change rebuild logic * fix some example * feat: check if k is larger than #docs * fix: WeChat history reader bugs and refactor wechat_rag to use unified architecture * fix email wrong -1 to process all file * refactor: reorgnize all examples/ and test/ * refactor: reorganize examples and add link checker * fix: add init.py * fix: handle certificate errors in link checker * fix wechat * merge * docs: update README to use proper module imports for apps - Change from 'python apps/xxx.py' to 'python -m apps.xxx' - More professional and pythonic module calling - Ensures proper module resolution and imports - Better separation between apps/ (production tools) and examples/ (demos) --------- Co-authored-by: yichuan520030910320 <yichuan_wang@berkeley.edu>	2025-08-03 23:06:24 -07:00
yichuan520030910320	19bcc07814	change readme discription	2025-07-28 20:52:45 -07:00
Andy Lee	261006c36a	docs: revert	2025-07-27 22:07:36 -07:00
yichuan520030910320	e9ee687472	nit: fix readme	2025-07-27 21:56:05 -07:00
yichuan520030910320	6f5d5e4a77	fix some readme	2025-07-27 21:50:09 -07:00
yichuan520030910320	e9d2d420bd	fix some readme	2025-07-27 20:48:23 -07:00
yichuan520030910320	ebabfad066	Merge branch 'main' of https://github.com/yichuan-w/LEANN	2025-07-27 20:44:36 -07:00
yichuan520030910320	e6f612b5e8	fix install and readme	2025-07-27 20:44:28 -07:00
Andy Lee	51c41acd82	docs: add comprehensive CONTRIBUTING.md guide with pre-commit setup	2025-07-27 20:40:42 -07:00
yichuan520030910320	455f93fb7c	fix emaple and add pypi example	2025-07-27 18:20:13 -07:00
yichuan520030910320	48207c3b69	add pypi example	2025-07-27 17:08:49 -07:00
yichuan520030910320	4de1caa40f	fix redame install method	2025-07-27 17:00:28 -07:00
yichuan520030910320	60eaa8165c	fix precommit and fix redame install method	2025-07-27 16:36:30 -07:00
yichuan520030910320	c1a5d0c624	fix readme	2025-07-27 02:24:28 -07:00
yichuan520030910320	383c6d8d7e	add clear instructions	2025-07-27 02:19:27 -07:00
yichuan520030910320	8596562de5	add pip install option to README	2025-07-27 02:06:40 -07:00
yichuan520030910320	7c8d7dc5c2	tones down	2025-07-26 21:47:55 -07:00
yichuan520030910320	0692bbf7a2	change workflow	2025-07-25 17:11:56 -07:00
yichuan520030910320	d91ce2e94d	readme	2025-07-25 02:19:54 -07:00
yichuan520030910320	d4f474c9b7	update broken link	2025-07-25 02:13:22 -07:00
yichuan520030910320	170f7644e9	simplify readme	2025-07-25 02:11:02 -07:00
yichuan520030910320	d038c81b8b	update benchmard section	2025-07-25 00:37:27 -07:00
yichuan520030910320	b6d43f5fd9	add gif	2025-07-25 00:12:35 -07:00
yichuan520030910320	5461b71d8c	colab dev	2025-07-24 19:03:46 -07:00
yichuan520030910320	efd6373b32	[chat] update huggingface chat and make qwen no thinking	2025-07-24 00:11:42 -07:00
yichuan520030910320	258a9a5c7f	[misc]test link again	2025-07-23 18:29:32 -07:00
yichuan520030910320	5d41ac6115	test link	2025-07-23 18:28:22 -07:00
yichuan520030910320	2a0fdb49b8	test link	2025-07-23 18:27:08 -07:00
yichuan520030910320	9d1b7231b6	fix broken link	2025-07-23 18:25:22 -07:00
yichuan520030910320	ed3095b478	fix broken link	2025-07-23 18:24:17 -07:00
yichuan520030910320	88eca75917	fix readme	2025-07-23 18:22:10 -07:00
yichuan520030910320	c083bda5b7	fix several bug	2025-07-23 18:17:11 -07:00
yichuan520030910320	99076e38bc	update install	2025-07-23 14:55:34 -07:00
yichuan520030910320	9698c1a02c	fix readme	2025-07-23 14:52:01 -07:00
yichuan520030910320	ae16d9d888	fix readme	2025-07-23 00:44:43 -07:00
yichuan520030910320	6e1af2eb0c	fix readme	2025-07-23 00:43:46 -07:00
yichuan520030910320	7695dd0d50	fix readme	2025-07-23 00:42:17 -07:00
yichuan520030910320	c2065473ad	fix readme	2025-07-23 00:30:42 -07:00
yichuan520030910320	c214b2e33e	fix readme	2025-07-23 00:09:24 -07:00
yichuan520030910320	f48f526f0a	fix readme	2025-07-22 23:21:15 -07:00

1 2 3

126 Commits