sij/khoj

mirror of https://github.com/khoj-ai/khoj.git synced 2024-12-18 02:27:10 +00:00

Author	SHA1	Message	Date
Debanjum Singh Solanky	c5ad172616	Keep loading animation at message end & reduce lists padding in Obsidian Previously loading animation would be at top of message. Moving it to bottom is more intuitve and easier to track. Remove white-space: pre from list elements. It was adding too much y axis padding to chat messages (and train of thought)	2024-07-23 17:56:03 +05:30
Debanjum Singh Solanky	54b4203683	Update chat API client tests to mix testing of batch and streaming mode	2024-07-23 17:56:03 +05:30
Debanjum Singh Solanky	3f5f418d0e	Use new chat streaming API to show Khoj train of thought in Obsidian client	2024-07-23 17:56:03 +05:30
Debanjum Singh Solanky	8303b09129	Convert snake case to camel case in chat view of obsidian plugin	2024-07-23 15:29:12 +05:30
Debanjum Singh Solanky	b224d7ffad	Simplify get_conversation_by_user DB adapter code	2024-07-23 14:51:11 +05:30
Debanjum Singh Solanky	daec439d52	Replace old chat router with new chat router with advanced streaming - Details Only return notes refs, online refs, inferred queries and generated response in non-streaming mode. Do not return train of throught and other status messages Incorporate missing logic from old chat API router into new one. - Motivation So we can halve chat API code by getting rid of the duplicate logic for the websocket router The deduplicated code: - Avoids inadvertant logic drift between the 2 routers - Improves dev velocity	2024-07-23 14:51:11 +05:30
Debanjum Singh Solanky	2d4b284218	Simplify streaming chat function in web client	2024-07-23 14:38:55 +05:30
Debanjum Singh Solanky	6b9550238f	Simplify advanced streaming chat API, align params with normal chat API	2024-07-22 22:51:24 +05:30
Debanjum Singh Solanky	b8d3e3669a	Stream Status Messages via Streaming Response from server to web client - Overview Use simpler HTTP Streaming Response to send status messages, alongside response and references from server to clients via API. Update web client to use the streamed response to show train of thought, stream response and render references. - Motivation This should allow other Khoj clients to pass auth headers and recieve Khoj's train of thought messages from server over simple HTTP streaming API. It'll also eventually deduplicate chat logic across /websocket and /chat API endpoints and help maintainability and dev velocity - Details - Pass references as a separate streaming message type for simpler parsing. Remove passing "### compiled references" altogether once the original /api/chat API is deprecated/merged with the new one and clients have been updated to consume the references using this new mechanism - Save message to conversation even if client disconnects. This is done by not breaking out of the async iterator that is sending the llm response. As the save conversation is called at the end of the iteration - Handle parsing chunked json responses as a valid json on client. This requires additional logic on client side but makes the client more robust to server chunking json response such that each chunk isn't itself necessarily a valid json.	2024-07-22 15:41:21 +05:30
Debanjum Singh Solanky	91fe41106e	Convert Websocket into Server Side Event (SSE) API endpoint - Convert functions in SSE API path into async generators using yields - Validate image generation, online, notes lookup and general paths of chat request are handled fine by the web client and server API	2024-07-21 14:20:22 +05:30
sabaimran	e694c82343	Fix Docker build issues with yarn / next /node (#859 ) * Rollback node version being installed from nodesource to node 20	2024-07-19 19:11:29 +05:30
sabaimran	1af9dbb083	Switch node/yarn install steps to use more native installation patterns	2024-07-19 17:10:08 +05:30
sabaimran	6d5ca5a3e1	yarn clean cache before build	2024-07-19 16:06:38 +05:30
sabaimran	7f0d1bd414	Add verbose logs when outputing yarn install steps	2024-07-19 15:48:43 +05:30
sabaimran	7426a4f819	Prefetch related agent when retrieving the conversation for performance improvements	2024-07-19 14:43:30 +05:30
Debanjum	2ab8fb78b1	Migrate the PyPI package to use project name: khoj (#853 ) ### Changes - Deprecate [khoj-assistant](https://pypi.org/project/khoj-assistant) pypi package. Use more accurate and succinct pypi project name, [khoj](https://pypi.org/project/khoj) - Update references to use `khoj` pypi package in docs and code - Update pypi workflow to publish to both khoj, khoj-assistant for now - Update stale python 3.9 support mentioned in our pyproject Can't support python 3.9 as depend on [Django 5.0.7](https://pypi.org/project/Django/5.0.7/) which needs python >=3.10 ### Verify - Updated `pypi.yml` github workflow publishes to both (new) [khoj](https://pypi.org/project/khoj/1.16.1.dev16/), (old) [khoj-assistant](https://pypi.org/project/khoj-assistant/1.16.1.dev16/) pypi projects - Can install Khoj python package with `pip install khoj`	2024-07-17 01:05:51 -07:00
Debanjum Singh Solanky	30d60aaae9	Add, fix Khoj Docker container labels	2024-07-17 10:41:17 +05:30
Debanjum Singh Solanky	583fa3c188	Migrate the pypi package to khoj project name. Update references - Deprecate khoj-assistant pypi package. Use more accurate and succinct pypi project name, khoj - Update references to sye khoj pypi package in docs and code instead of the legacy khoj-assistant pypi package - Update pypi workflow to publish to both khoj, khoj-assistant for now - Update stale python 3.9 support mentioned in our pyproject. Can't support python 3.9 as depend on latest django which support >=3.10	2024-07-17 10:41:16 +05:30
Debanjum	23f61d49e0	Support syncing, searching images from Obsidian plugin (#847 ) - Sync images from Obsidian vault with Khoj server now that Khoj can OCR images - Support rendering images returned by Khoj search modal	2024-07-14 20:41:39 -07:00
Debanjum Singh Solanky	02658ad4fd	Upgrade Django version	2024-07-11 16:35:10 +05:30
Debanjum Singh Solanky	cbae8b68fb	Add DB migration from making bi_encode configs optional in #834	2024-07-11 16:33:31 +05:30
Debanjum Singh Solanky	3a75838196	Add Keyboard shortcuts to navigate in Khoj Desktop	2024-07-11 16:29:53 +05:30
Debanjum Singh Solanky	6c1861b319	Improve the prompt to generate images with DALLE3 and SD3 - Major - Ask for prompt in prose - Remove seed from SD3 image generation to improve diversity of output for a given prompt Otherwise for conversations with similar sounding prompts, the images would be almost exactly the same. This maybe another indicator of SD3's inability to capture detailed instructions - Consistently use "prompt" wording instead of "query" in improved image generation prompts. Previously a mix of those terms were being used, which could confuse the chat model - Minor - Add day of week to prompt - Remove 2-5 sentence limit on instructions to SD3. It seems to be able to follow longer instructions just with less fidelity than DALLE. And the 2-5 sentence instruction limit wasn't being adhered to - Improve ability to edit, improve the image based on follow-up instructions by the user - Align prompts for DALLE and SD3. Only difference is to wrap text to be rendered in quotes for SD3. This improves it's ability to render requested text. DALLE cannot render text as well or consistently	2024-07-11 16:29:53 +05:30
Debanjum Singh Solanky	21fe1a917b	Support syncing, searching images from Obsidian plugin	2024-07-11 16:22:31 +05:30
sabaimran	260aa61818	Remove tests for python3.9	2024-07-09 12:28:11 +05:30
sabaimran	4471c1e37f	Apply mitigations for piling up open connections - Because we're using a FastAPI api framework with a Django ORM, we're running into some interesting conditions around connection pooling and clean-up. We're ending up with a large pile-up of open, stale connections to the DB recurringly when the server has been running for a while. To mitigate this problem, given starlette and django run in different python threads, add a middleware that will go and call the connection clean up method in each of the threads.	2024-07-09 12:22:58 +05:30
Debanjum	0b1b262512	Add system dependencies required by RapidOCR to fix Khoj Docker image (#842 ) - Issue The Khoj docker build would fail with `ImportError: libGL.so.1: cannot open shared object file: No such file or directory`. This was required by the Khoj RapidOCR python package dependency. - Fix A minimal set of system packages have been added to resolve this issue.	2024-07-08 22:16:16 +05:30
kxnarak	43413cd21f	add dependencies required by the RapidOCR python package	2024-07-08 18:26:19 +05:30
sabaimran	037e157648	Fix a variety of links	2024-07-08 16:49:13 +05:30
sabaimran	6b80bb3f37	Add a demo for the khoj mini application, minor updates to other pages, remove out of date demos page	2024-07-08 16:33:47 +05:30
Debanjum Singh Solanky	9e31ebff93	Release Khoj version 1.16.0	2024-07-07 18:26:10 +05:30
Debanjum Singh Solanky	54132efd67	Fix Khoj Obsidian plugin build	2024-07-07 18:26:10 +05:30
Debanjum Singh Solanky	510d9b3a29	Add short keys to open chat menu, new chat, search from Obsidian pane	2024-07-07 17:57:17 +05:30
Debanjum Singh Solanky	3e0c882e27	Transcribe only when keyboard shortcut or button pressed in Obsidian - Transcribe on holding Ctrl+s keyboard shortcut - Transcribe on holding the transcribe button pressed via mouse too - Make the transcribe button robust to inadvertent touches by using timeout - Do not transcribe, trigger auto-send on silences. Silence detection is super rudimentary, just blocks standard emanations by whisper when no speech	2024-07-07 17:57:17 +05:30
sabaimran	0eb000c3ea	Add health checks for the django ORM	2024-07-07 16:11:28 +05:30
Debanjum Singh Solanky	a31cd0dec1	Fix async batch delete of indexed entries	2024-07-06 22:45:26 +05:30
Debanjum	08b379c2ab	Fix, Improve Indexing, Deleting Files (#840 ) ### Fix - Fix degrade in speed when indexing large files - Resolve org-mode indexing bug by splitting current section only once by heading - Improve summarization by fixing formatting of text in indexed files ### Improve - Improve scaling user, admin flows to delete all entries for a user	2024-07-06 19:52:42 +05:30
Debanjum Singh Solanky	4a471979eb	Upgrade sentence-transformer package to version 3.0.1 Add einops dependency for some sentence transformer models like the nomic-embed	2024-07-06 19:35:59 +05:30
Debanjum Singh Solanky	d693baccbc	Make it optional to set the encoder, cross-encoder configs via admin UI	2024-07-06 19:35:59 +05:30
Debanjum Singh Solanky	1baebb8d0e	Identify markdown headings by any whitespace character after ^#+ Previously only markdown headings with space characters after # would be considered a heading. So ^##\t wouldn't be considered a valid heading	2024-07-06 19:35:59 +05:30
Debanjum Singh Solanky	010486fb36	Split current section once by heading to resolve org-mode indexing bug - Split once by heading (=first_non_empty) to extract current section body Otherwise child headings with same prefix as current heading will cause the section split to go into infinite loop - Also add check to prevent getting into recursive loop while trying to split entry into sub sections	2024-07-06 19:35:59 +05:30
Debanjum Singh Solanky	6a135b1ed7	Fix degrade in speed of indexing large files. Improve summarization Adding files to the DB for summarization was slow, buggy in two ways: - We were updating same text of modified files in DB = no of chunks per file times - The `" ".join(file_content)' code was breaking each character in the file content by a space. This formats the original file content incorrectly before storing in the DB Because this code ran in the main file indexing path, it was slowing down file indexing. Knowledge bases with larger files were impacted more strongly	2024-07-06 19:35:59 +05:30
Debanjum Singh Solanky	e6ffb6b52c	Improve scaling user flow to delete all entries - Delete entries by batch to improve efficiency of query at scale - Share code to delete all user entries between it's async, sync methods - Add indicator to show when files being deleted on web config page	2024-07-06 19:35:59 +05:30
Debanjum Singh Solanky	1ab59865b5	Improve scaling admin flow to delete all entries for user	2024-07-06 19:35:59 +05:30
Debanjum	05138cbd0a	Use DOM Scripting, Add CSP to Web config pages. Disable CSP in Obsidian plugin (#834 ) - Add CSP to web config pages. Load phone no. validation js, css from S3 - Construct config page elements on Web via DOM scripting - Disable CSP in Khoj Obsidian as it interferes with Obsidian functionality - Other miscellaneous voice message level improvements (rate limit, listening animation)	2024-07-06 19:30:09 +05:30
Debanjum Singh Solanky	9bdb48807b	Ratelimit text to speech model. Validate share chat url domain - Do not log auth error message on server when Resend setup as Magic links for sign-in are now supported	2024-07-06 12:53:19 +05:30
Debanjum Singh Solanky	b334db0fca	Add CSP to web config pages. Load phone no validation js, css from S3	2024-07-06 12:48:28 +05:30
Debanjum Singh Solanky	2f034f807a	Construct config page elements on Web via DOM scripting. Minimize isage of innerHTML to prevent DOM clobbering and unintended escape by user Input	2024-07-06 12:48:28 +05:30
Debanjum Singh Solanky	69c9e8cc08	Disable CSP in Khoj Obsidian as it interferes with Obsidian functionality The Khoj CSP interferes with other Obsidian features and plugins as CSP is applied page wide. For now chat message sanitization via Dompurify should suffice. Enable CSP when can scope it to only the Khoj Obsidian plugin.	2024-07-05 16:10:08 +05:30
Debanjum Singh Solanky	a353d883a0	Make it optional to set the encoder, cross-encoder configs via admin UI Upgrade sentence-transformer, add einops dependency for some sentence transformer models like nomic	2024-07-05 16:09:30 +05:30

1 2 3 4 5 ...

2996 commits