削微寒
|
0bdb9a32f2
|
fix HelloGitHub Badge code (#313)
|
2024-09-22 16:31:12 +07:00 |
|
Khoi-Nguyen Nguyen-Ngoc
|
a865e2b095
|
feat: modify base dependencies + remove unnecessary packages in lite docker (#310)
* feat: update base/adv dependencies
* feat: update Dockerfile
* ci: update free disk for docker build
|
2024-09-21 12:11:58 +07:00 |
|
Quang (Albert)
|
d6a9510441
|
fix: turn off commitlint job (#304)
|
2024-09-18 09:48:54 +07:00 |
|
Quang (Albert)
|
7762190d05
|
feat: add local theme (#288)
* feat: add local theme instead of from hub
* chore: add credit
* fix: typo
|
2024-09-17 19:03:39 +07:00 |
|
Anush
|
e2bd78e9c4
|
feat: Qdrant vectorstore support (#260)
* feat: Qdrant vectorstore support
* chore: review changes
* docs: Updated README.md
|
2024-09-16 04:17:36 +07:00 |
|
Tadashi
|
cbe45a4395
|
docs: update README
|
2024-09-13 10:55:16 +07:00 |
|
Tadashi
|
463890745c
|
docs: update README
|
2024-09-12 21:23:12 +07:00 |
|
kan_cin
|
d3fd75297f
|
feat: add multi-stages docker and support platform arm (#274)
* feat: add multi-stages docker and support platform arm
* refactor: pre-commit
* fix: raise ImportError (fastembed) instead of auto install
* feat: add dependencies for local llm
* feat: free disk
* feat: update README
* feat: update README
* chore: fix typo
---------
Co-authored-by: cin-niko <niko@cinnamon.is>
|
2024-09-12 20:25:03 +07:00 |
|
mst
|
73a476979e
|
fix: change column type to string for relation_type (#272) #none
|
2024-09-11 20:47:03 +07:00 |
|
Tadashi
|
cd85c4935c
|
docs: update badge #none
|
2024-09-10 15:28:26 +07:00 |
|
Tuan Anh Nguyen Dang (Tadashi_Cin)
|
96d2086017
|
fix: add guidance parameters for LC wrapper models (#255)
* fix: add docstring to LC wrapper models
* fix: fix metadata passing with LC embedding wrapper
|
2024-09-09 14:15:34 +07:00 |
|
Tadashi
|
ce489725d8
|
ci: revert GH env var
|
2024-09-08 21:39:16 +07:00 |
|
Tadashi
|
9bfb5ef778
|
docs: fix typos
|
2024-09-08 21:31:05 +07:00 |
|
Tadashi
|
2d6c02ebea
|
fix: update README bump:patch
|
2024-09-08 21:29:40 +07:00 |
|
Tuan Anh Nguyen Dang (Tadashi_Cin)
|
b06c4777a3
|
fix: add PDFJS download to Windows setup (#249)
|
2024-09-08 21:22:01 +07:00 |
|
kan_cin
|
855f3df75f
|
fix: change github token (#235) #none
|
2024-09-08 10:55:45 +07:00 |
|
kan_cin
|
dbb6bb275f
|
feat: add test connection for edit spec (#239)
|
2024-09-08 10:55:13 +07:00 |
|
Quang (Albert)
|
fa881d4450
|
feat: add Portable Git to Windows installer (#232)
* feat(windows installer): check and install git
* feat: update run_windows.bat
* feat: Replace standalone Git installer with Portable Git
* feat: support milvus vector db (#188) #none
Signed-off-by: ChengZi <chen.zhang@zilliz.com>
* feat: add github action to build docker for release (#168) #none
* feat: update build push docker action
* feat: remove tag trigger
* feat: remove manual trigger
* fix: update workflow
* feat: update build-push-docker.yaml
* fix: update workflow
* fix: update workflow
* fix: update workflow
* refactor: comfort pre-commit
* feat: update permission
* feat: update docker support pdfjs
* refactor: comfort pre-commit
* feat: add support for Gemini, Claude through Langchain (#225) (bump:patch)
* fix: disable default install for google-genai package
* fix: disable default install for anthropic
* fix: update on release event build push docker (#228) #none
* fix: update on release event build push docker
* refactor: comfort pre-commit
* fix: limit fastapi version (#229)
* fix: update requirements (#230)
* style: fix pre-commit
---------
Signed-off-by: ChengZi <chen.zhang@zilliz.com>
Co-authored-by: ChengZi <chen.zhang@zilliz.com>
Co-authored-by: kan_cin <kan@cinnamon.is>
Co-authored-by: Tuan Anh Nguyen Dang (Tadashi_Cin) <tadashi@cinnamon.is>
|
2024-09-08 10:54:26 +07:00 |
|
Tuan Anh Nguyen Dang (Tadashi_Cin)
|
069f0f3c83
|
feat: expose Cohere and HF embedding support on UI (#236)
|
2024-09-06 18:18:19 +07:00 |
|
taprosoft
|
4d7f16475f
|
docs: update default Docker image instruction
|
2024-09-06 03:07:42 +00:00 |
|
Tuan Anh Nguyen Dang (Tadashi_Cin)
|
ef7e91fcae
|
fix: update requirements (#230)
|
2024-09-06 09:36:21 +07:00 |
|
Tuan Anh Nguyen Dang (Tadashi_Cin)
|
e2ed3564ce
|
fix: limit fastapi version (#229)
|
2024-09-06 09:23:26 +07:00 |
|
kan_cin
|
4b0b28227d
|
fix: update on release event build push docker (#228) #none
* fix: update on release event build push docker
* refactor: comfort pre-commit
|
2024-09-05 23:44:54 +07:00 |
|
Tadashi
|
318895b287
|
fix: disable default install for anthropic
|
2024-09-05 23:18:53 +07:00 |
|
Tadashi
|
3267e6c654
|
fix: disable default install for google-genai package
|
2024-09-05 23:08:28 +07:00 |
|
Tuan Anh Nguyen Dang (Tadashi_Cin)
|
05245f501c
|
feat: add support for Gemini, Claude through Langchain (#225) (bump:patch)
|
2024-09-05 21:58:20 +07:00 |
|
kan_cin
|
8be8a4a9d0
|
feat: add github action to build docker for release (#168) #none
* feat: update build push docker action
* feat: remove tag trigger
* feat: remove manual trigger
* fix: update workflow
* feat: update build-push-docker.yaml
* fix: update workflow
* fix: update workflow
* fix: update workflow
* refactor: comfort pre-commit
* feat: update permission
* feat: update docker support pdfjs
* refactor: comfort pre-commit
|
2024-09-05 15:02:23 +07:00 |
|
ChengZi
|
772186b6e5
|
feat: support milvus vector db (#188) #none
Signed-off-by: ChengZi <chen.zhang@zilliz.com>
|
2024-09-04 20:22:50 +07:00 |
|
Tuan Anh Nguyen Dang (Tadashi_Cin)
|
76f2652d2a
|
fix: re-enable tests and fix legacy test interface (#208)
* fix: re-enable tests and fix legacy test interface
* fix: skip llamacpp based on installed status
* fix: minor fix
|
2024-09-04 12:37:39 +07:00 |
|
Tadashi
|
92f6b8e1bf
|
fix: update README (bump:patch)
|
2024-09-04 08:05:21 +07:00 |
|
Tuan Anh Nguyen Dang (Tadashi_Cin)
|
607867d7e6
|
feat: add markdown file support (#202)
* feat: add support for .md
* fix: disable download all on private collection
|
2024-09-03 23:15:26 +07:00 |
|
nguyen
|
4f0785773d
|
feat: reduce docker image size by removing unnecessary cache (#174) (#none)
* feat: reduce docker image size by removing unnecessary cache
* fix: trailing whitespace
|
2024-09-02 18:13:53 +07:00 |
|
Tuan Anh Nguyen Dang (Tadashi_Cin)
|
35b2927e5c
|
fix: update app version resolver in flowsettings (#180) (bump:patch)
|
2024-09-02 17:42:39 +07:00 |
|
Le Minh Duc
|
4d5f9ba39c
|
ci: add commitlint (#170)
|
2024-09-01 23:10:03 +07:00 |
|
kan_cin
|
041d229282
|
feat: add test connection feature (#166)
* feat: add test connection feature
* fix: typo
|
2024-09-01 08:22:36 +07:00 |
|
Tadashi
|
c1e8c37e5e
|
fix: update packaging script (bump:patch)
|
2024-08-31 07:07:28 +07:00 |
|
Tadashi
|
7daa9eb149
|
docs: update demo URL (bump:minor)
|
2024-08-30 23:46:18 +07:00 |
|
Tuan Anh Nguyen Dang (Tadashi_Cin)
|
09f8f91510
|
docs: update README (#157)
|
2024-08-30 23:29:31 +07:00 |
|
Tuan Anh Nguyen Dang (Tadashi_Cin)
|
9354ad8241
|
fix: update default settings and local model guide (#156)
|
2024-08-30 23:18:31 +07:00 |
|
Quang (Albert)
|
4b2b334d2c
|
fix: refine kotaemon/pyproject.toml (#153)
|
2024-08-30 23:02:14 +07:00 |
|
Tuan Anh Nguyen Dang (Tadashi_Cin)
|
d880294153
|
fix: pwd change in setttings (#147)
|
2024-08-29 13:41:12 +07:00 |
|
ian
|
971ffcc9d0
|
add github star history (#137)
|
2024-08-28 17:19:20 +07:00 |
|
Quang (Albert)
|
fcefb80fa6
|
feat: Add contribution templates (#none) (#139)
* feat: Add PR template
* feat: Add issue templates
* style: Comfort pre-commit
* style: Comfort pre-commit
|
2024-08-28 17:18:50 +07:00 |
|
John Freier
|
1cdefe7ba3
|
Update mkdocs.yml (#129)
Documentation Navigation URL Fix
|
2024-08-28 06:37:17 +07:00 |
|
ian
|
5946fd33de
|
change default bump to patch, don't create release if there is no bump (#126)
|
2024-08-28 06:30:53 +07:00 |
|
Tuan Anh Nguyen Dang (Tadashi_Cin)
|
bb56ef4f8e
|
chore: update workflow (#124)
|
2024-08-26 09:52:16 +07:00 |
|
Tuan Anh Nguyen Dang (Tadashi_Cin)
|
2570e11501
|
feat: merge develop (#123)
* Support hybrid vector retrieval
* Enable figures and table reading in Azure DI
* Retrieve with multi-modal
* Fix mixing up table
* Add txt loader
* Add Anthropic Chat
* Raising error when retrieving help file
* Allow same filename for different people if private is True
* Allow declaring extra LLM vendors
* Show chunks on the File page
* Allow elasticsearch to get more docs
* Fix Cohere response (#86)
* Fix Cohere response
* Remove Adobe pdfservice from dependency
kotaemon doesn't rely more pdfservice for its core functionality,
and pdfservice uses very out-dated dependency that causes conflict.
---------
Co-authored-by: trducng <trungduc1992@gmail.com>
* Add confidence score (#87)
* Save question answering data as a log file
* Save the original information besides the rewritten info
* Export Cohere relevance score as confidence score
* Fix style check
* Upgrade the confidence score appearance (#90)
* Highlight the relevance score
* Round relevance score. Get key from config instead of env
* Cohere return all scores
* Display relevance score for image
* Remove columns and rows in Excel loader which contains all NaN (#91)
* remove columns and rows which contains all NaN
* back to multiple joiner options
* Fix style
---------
Co-authored-by: linhnguyen-cinnamon <cinmc0019@CINMC0019-LinhNguyen.local>
Co-authored-by: trducng <trungduc1992@gmail.com>
* Track retriever state
* Bump llama-index version 0.10
* feat/save-azuredi-mhtml-to-markdown (#93)
* feat/save-azuredi-mhtml-to-markdown
* fix: replace os.path to pathlib change theflow.settings
* refactor: base on pre-commit
* chore: move the func of saving content markdown above removed_spans
---------
Co-authored-by: jacky0218 <jacky0218@github.com>
* fix: losing first chunk (#94)
* fix: losing first chunk.
* fix: update the method of preventing losing chunks
---------
Co-authored-by: jacky0218 <jacky0218@github.com>
* fix: adding the base64 image in markdown (#95)
* feat: more chunk info on UI
* fix: error when reindexing files
* refactor: allow more information exception trace when using gpt4v
* feat: add excel reader that treats each worksheet as a document
* Persist loader information when indexing file
* feat: allow hiding unneeded setting panels
* feat: allow specific timezone when creating conversation
* feat: add more confidence score (#96)
* Allow a list of rerankers
* Export llm reranking score instead of filter with boolean
* Get logprobs from LLMs
* Rename cohere reranking score
* Call 2 rerankers at once
* Run QA pipeline for each chunk to get qa_score
* Display more relevance scores
* Define another LLMScoring instead of editing the original one
* Export logprobs instead of probs
* Call LLMScoring
* Get qa_score only in the final answer
* feat: replace text length with token in file list
* ui: show index name instead of id in the settings
* feat(ai): restrict the vision temperature
* fix(ui): remove the misleading message about non-retrieved evidences
* feat(ui): show the reasoning name and description in the reasoning setting page
* feat(ui): show version on the main windows
* feat(ui): show default llm name in the setting page
* fix(conf): append the result of doc in llm_scoring (#97)
* fix: constraint maximum number of images
* feat(ui): allow filter file by name in file list page
* Fix exceeding token length error for OpenAI embeddings by chunking then averaging (#99)
* Average embeddings in case the text exceeds max size
* Add docstring
* fix: Allow empty string when calling embedding
* fix: update trulens LLM ranking score for retrieval confidence, improve citation (#98)
* Round when displaying not by default
* Add LLMTrulens reranking model
* Use llmtrulensscoring in pipeline
* fix: update UI display for trulen score
---------
Co-authored-by: taprosoft <tadashi@cinnamon.is>
* feat: add question decomposition & few-shot rewrite pipeline (#89)
* Create few-shot query-rewriting. Run and display the result in info_panel
* Fix style check
* Put the functions to separate modules
* Add zero-shot question decomposition
* Fix fewshot rewriting
* Add default few-shot examples
* Fix decompose question
* Fix importing rewriting pipelines
* fix: update decompose logic in fullQA pipeline
---------
Co-authored-by: taprosoft <tadashi@cinnamon.is>
* fix: add encoding utf-8 when save temporal markdown in vectorIndex (#101)
* fix: improve retrieval pipeline and relevant score display (#102)
* fix: improve retrieval pipeline by extending first round top_k with multiplier
* fix: minor fix
* feat: improve UI default settings and add quick switch option for pipeline
* fix: improve agent logics (#103)
* fix: improve agent progres display
* fix: update retrieval logic
* fix: UI display
* fix: less verbose debug log
* feat: add warning message for low confidence
* fix: LLM scoring enabled by default
* fix: minor update logics
* fix: hotfix image citation
* feat: update docx loader for handle merged table cells + handle zip file upload (#104)
* feat: update docx loader for handle merged table cells
* feat: handle zip file
* refactor: pre-commit
* fix: escape text in download UI
* feat: optimize vector store query db (#105)
* feat: optimize vector store query db
* feat: add file_id to chroma metadatas
* feat: remove unnecessary logs and update migrate script
* feat: iterate through file index
* fix: remove unused code
---------
Co-authored-by: taprosoft <tadashi@cinnamon.is>
* fix: add openai embedidng exponential back-off
* fix: update import download_loader
* refactor: codespell
* fix: update some default settings
* fix: update installation instruction
* fix: default chunk length in simple QA
* feat: add share converstation feature and enable retrieval history (#108)
* feat: add share converstation feature and enable retrieval history
* fix: update share conversation UI
---------
Co-authored-by: taprosoft <tadashi@cinnamon.is>
* fix: allow exponential backoff for failed OCR call (#109)
* fix: update default prompt when no retrieval is used
* fix: create embedding for long image chunks
* fix: add exception handling for additional table retriever
* fix: clean conversation & file selection UI
* fix: elastic search with empty doc_ids
* feat: add thumbnail PDF reader for quick multimodal QA
* feat: add thumbnail handling logic in indexing
* fix: UI text update
* fix: PDF thumb loader page number logic
* feat: add quick indexing pipeline and update UI
* feat: add conv name suggestion
* fix: minor UI change
* feat: citation in thread
* fix: add conv name suggestion in regen
* chore: add assets for usage doc
* chore: update usage doc
* feat: pdf viewer (#110)
* feat: update pdfviewer
* feat: update missing files
* fix: update rendering logic of infor panel
* fix: improve thumbnail retrieval logic
* fix: update PDF evidence rendering logic
* fix: remove pdfjs built dist
* fix: reduce thumbnail evidence count
* chore: update gitignore
* fix: add js event on chat msg select
* fix: update css for viewer
* fix: add env var for PDFJS prebuilt
* fix: move language setting to reasoning utils
---------
Co-authored-by: phv2312 <kat87yb@gmail.com>
Co-authored-by: trducng <trungduc1992@gmail.com>
* feat: graph rag (#116)
* fix: reload server when add/delete index
* fix: rework indexing pipeline to be able to disable vectorstore and splitter if needed
* feat: add graphRAG index with plot view
* fix: update requirement for graphRAG and lighten unnecessary packages
* feat: add knowledge network index (#118)
* feat: add Knowledge Network index
* fix: update reader mode setting for knet
* fix: update init knet
* fix: update collection name to index pipeline
* fix: missing req
---------
Co-authored-by: jeff52415 <jeff.yang@cinnamon.is>
* fix: update info panel return for graphrag
* fix: retriever setting graphrag
* feat: local llm settings (#122)
* feat: expose context length as reasoning setting to better fit local models
* fix: update context length setting for agents
* fix: rework threadpool llm call
* fix: fix improve indexing logic
* fix: fix improve UI
* feat: add lancedb
* fix: improve lancedb logic
* feat: add lancedb vectorstore
* fix: lighten requirement
* fix: improve lanceDB vs
* fix: improve UI
* fix: openai retry
* fix: update reqs
* fix: update launch command
* feat: update Dockerfile
* feat: add plot history
* fix: update default config
* fix: remove verbose print
* fix: update default setting
* fix: update gradio plot return
* fix: default gradio tmp
* fix: improve lancedb docstore
* fix: fix question decompose pipeline
* feat: add multimodal reader in UI
* fix: udpate docs
* fix: update default settings & docker build
* fix: update app startup
* chore: update documentation
* chore: update README
* chore: update README
---------
Co-authored-by: trducng <trungduc1992@gmail.com>
* chore: update README
* chore: update README
---------
Co-authored-by: trducng <trungduc1992@gmail.com>
Co-authored-by: cin-ace <ace@cinnamon.is>
Co-authored-by: Linh Nguyen <70562198+linhnguyen-cinnamon@users.noreply.github.com>
Co-authored-by: linhnguyen-cinnamon <cinmc0019@CINMC0019-LinhNguyen.local>
Co-authored-by: cin-jacky <101088014+jacky0218@users.noreply.github.com>
Co-authored-by: jacky0218 <jacky0218@github.com>
Co-authored-by: kan_cin <kan@cinnamon.is>
Co-authored-by: phv2312 <kat87yb@gmail.com>
Co-authored-by: jeff52415 <jeff.yang@cinnamon.is>
|
2024-08-26 08:50:37 +07:00 |
|
ian
|
86d60e1649
|
Update docs (#88)
Co-authored-by: ian <ian@cinnamon.is>
|
2024-05-31 17:49:02 +07:00 |
|
trducng
|
ebf1315569
|
(pump:minor) Allow the indexing pipeline to report the indexing progress onto the UI (#81)
* Turn the file indexing event to generator to report progress
* Fix React text's trimming function
* Refactor delete file into a method
|
2024-05-25 22:09:41 +07:00 |
|
trducng
|
56dfc8fb53
|
Allow the application name to be configurable in settings (#80)
* Make app name configurable
* Use app name in browser tab
|
2024-05-20 22:37:24 +07:00 |
|