sij/khoj

mirror of https://github.com/khoj-ai/khoj.git synced 2024-11-23 23:48:56 +01:00

Author	SHA1	Message	Date
Debanjum	1a83bbcc94	Clean API chat router. Move FeedbackData response type to router helper	2024-11-01 17:51:41 -07:00
sabaimran	e6eb87bbb5	Merge branch 'improve-debug-reasoning-and-other-misc-fixes' of github.com:khoj-ai/khoj into improve-debug-reasoning-and-other-misc-fixes	2024-11-01 16:48:39 -07:00
sabaimran	a213b593e8	Limit the number of urls the webscraper can extract for scraping	2024-11-01 16:48:36 -07:00
sabaimran	327fcb8f62	create defiltered query after conversation command is extracted	2024-11-01 16:48:03 -07:00
sabaimran	b79a9ec36d	Clarify description of the code evaluation environment: not for document creation	2024-11-01 16:47:27 -07:00
Debanjum	9c7b36dc69	Use standard per minute rate limits across user types	2024-11-01 16:16:06 -07:00
Debanjum	ac21b10dd5	Simplify logic to get default search model. Remove unused import	2024-11-01 15:14:00 -07:00
sabaimran	2b35790165	Merge branch 'master' of github.com:khoj-ai/khoj into improve-debug-reasoning-and-other-misc-fixes	2024-11-01 14:51:26 -07:00
Debanjum	22f3ed3f5d	Research Mode: Give Khoj the ability to perform more advanced reasoning (#952 ) ## Overview Khoj can now go into research mode and use a python code interpreter. These are experimental features that are being released early for feedback and testing. - Research mode allows Khoj to dynamically select the tools it needs to best answer the question. It is also allowed more iterations to get to a satisfactory answer. Its more dynamic train of thought is shown to improve visibility into its thinking. - Adding ability for Khoj to use a python code interpreter is an adjacent capability. It can help Khoj do some data analysis and generate charts for you. A sandboxed python to run code is provided using [cohere-terrarium](https://github.com/cohere-ai/cohere-terrarium?tab=readme-ov-file), [pyodide](https://pyodide.org/). ## Analysis Research mode (significantly?) improves Khoj's information retrieval for more complex queries requiring multi-step lookups but takes longer to run. It can research for longer, requiring less back-n-forth with the user to find an answer. Research mode gives most gains when used with more advanced chat models (like o1, 4o, new claude sonnet and gemini-pro-002). Smaller models improve their response quality but tend to get into repetitive loops more often. ## Next Steps - Get community feedback on research mode. What works, what fails, what is confusing, what'd be cool to have. - Tune Khoj's capabilities for longer autonomous runs and to generalize across a larger range of model sizes ## Miscellaneous Improvements - Khoj's train of thought is saved and shown for all messages, not just the latest one - Render charts generated by Khoj and code running using the code tool on the web app - Align chat input color to currently selected agent color	2024-11-01 14:46:29 -07:00
sabaimran	baa939f4ce	When running code, strip any code delimiters. Disable application json type specification in Gemini request.	2024-11-01 13:47:39 -07:00
sabaimran	8fd2fe162f	Determine if research mode is enabled by checking the conversation commands and 'linting' them in the selection phase	2024-11-01 13:12:34 -07:00
sabaimran	cead1598b9	Don't reset research mode after completing research execution	2024-11-01 13:00:11 -07:00
Debanjum	c1c779a7ef	Do not yaml format raw code results in context for LLM. It's confusing	2024-11-01 12:45:26 -07:00
sabaimran	b3dad1f393	Standardize rate limits to 1/6 ratio	2024-11-01 12:21:09 -07:00
sabaimran	23a49b6b95	Add documentation for python code execution capability	2024-11-01 12:14:33 -07:00
Debanjum	cd75151431	Do not allow auto selecting research mode as tool for now. You are required to manually turning it on. This takes longer and should be a high intent activity initiated by user	2024-11-01 12:07:52 -07:00
Debanjum	0b0cfb35e6	Simplify in research mode check in api_chat. - Dedent code for readability - Use better name for in research mode check - Continue to remove inferred summarize command when multiple files in file filter even when not in research mode - Continue to show select information source train of thought. It was removed by mistake earlier	2024-11-01 12:07:08 -07:00
sabaimran	ffa7f95559	Add template for a code sandbox to the docker-compose configuration	2024-11-01 11:50:58 -07:00
Debanjum	73750ef286	Merge branch 'master' into features/advanced-reasoning	2024-11-01 11:42:01 -07:00
sabaimran	1fc280db35	Handle case where infer_webpage_url returns no valid urls	2024-11-01 11:41:32 -07:00
Debanjum	1c920273dd	Add Prompt Tracer to Visualize, Analyze and Debug Khoj's Train of Thought (#951 ) ## Overview Use git to capture prompt traces of khoj's train of thought. View, analyze and debug them using your favorite git client (e.g vscode, magit). - Each commit captures an interaction with an LLM The commit writes the query, response and system message each to a separate file in the repo. The commit message captures the chat model, Khoj version and other metadata - Each conversation turn can have multiple interactions with an LLM (e.g Khoj's train of thought) - Each new conversation turn forks from and merges back into its conversation branch - Each new conversation branches from the user branch - Each new user branches from root commit on the main branch ## Usage 1. Set `KHOJ_DEBUG=true` or start khoj in very verbose mode with `khoj -vv` to turn on prompt tracing 2. Chat with Khoj as usual 3. Open the promptrace git repo to view the generated prompt traces using your favorite git porcelain. The Khoj prompt trace git repo is created at `/tmp/khoj_promptrace` by default. You can configure the prompt trace directory by setting the `PROMPTRACE_DIR`environment variable. ## Implementation - Add utility functions to capture prompt traces using git (via `gitpython`) - Make each model provider in Khoj commit their LLM interactions with promptrace - Weave chat metadata from chat API through all chat actors and commit it to the prompt trace	2024-11-01 11:33:54 -07:00
sabaimran	33d36ee58c	Add experimental notice to research mode tooltip	2024-11-01 11:00:27 -07:00
sabaimran	0145b2a366	Set usage limits on the research mode	2024-11-01 10:29:33 -07:00
sabaimran	3ea94ac972	Only include inferred-queries in chat history when present	2024-10-31 22:01:41 -07:00
sabaimran	149cbe1019	Use bottom anchor for the commandbar popover	2024-10-31 20:40:38 -07:00
sabaimran	21858acccc	Remove conversation command always in query, filter out inferred queries that were not with selected tool when going through tool selection iterations	2024-10-31 20:27:38 -07:00
sabaimran	19241805ee	Merge branch 'master' of github.com:khoj-ai/khoj into improve-debug-reasoning-and-other-misc-fixes	2024-10-31 18:20:23 -07:00
Debanjum	302bd51d17	Improve online chat actor prompt for research and normal mode - Match the online query generator prompt to match the formatting of extract questions - Separate iteration results by newline - Improve webpage and online tool descriptions	2024-10-31 18:17:12 -07:00
Debanjum	52163fe299	Improve research planner prompt to reduce looping	2024-10-31 18:17:01 -07:00
sabaimran	7ebf999688	Merge branch 'master' of github.com:khoj-ai/khoj into features/advanced-reasoning	2024-10-31 18:15:13 -07:00
sabaimran	159ea44883	Remove frame references in the diagramming prompts	2024-10-31 18:14:51 -07:00
Debanjum	89597aefe9	Json dump contents in prompt tracer to make structure discernable	2024-10-31 18:08:42 -07:00
Debanjum	5b15176e20	Only add /research prefix in research mode if not already in user query	2024-10-31 18:08:42 -07:00
sabaimran	559601dd0a	Do not exit if/else loop in research loop when notes not found	2024-10-31 13:51:10 -07:00
sabaimran	a13760640c	Only show trash can when turnId is present	2024-10-31 13:19:16 -07:00
sabaimran	8d1ecb9bd8	Add optional brew steps for docker install	2024-10-31 12:41:53 -07:00
Debanjum	adca6cbe9d	Merge branch 'master' into add-prompt-tracer-for-observability	2024-10-31 02:28:34 -07:00
Debanjum	e17dc9f7b5	Put train of thought ui before Khoj response on web app	2024-10-31 02:24:53 -07:00
Debanjum	e8e6ead39f	Fix deleting new messages generated after conversation load	2024-10-30 20:56:38 -07:00
Debanjum	cb90abc660	Resolve train of thought component needs unique key id error on web app	2024-10-30 14:00:21 -07:00
Debanjum	ca5a6831b6	Add ability to delete messages from the web app	2024-10-30 14:00:21 -07:00
Debanjum	ba15686682	Store turn id with each chat message. Expose API to delete chat turn Each chat turn is a user query, khoj response message pair	2024-10-30 14:00:21 -07:00
Debanjum	f64f5b3b6e	Handle add/delete file filter operation on non-existent conversation	2024-10-30 14:00:21 -07:00
Debanjum	b3a63017b5	Support setting seed for reproducible LLM response generation Anthropic models do not support seed. But offline, gemini and openai models do. Use these to debug and test Khoj via KHOJ_LLM_SEED env var	2024-10-30 14:00:21 -07:00
Debanjum	d44e68ba01	Improve handling embedding model config from admin interface - Allow server to start if loading embedding model fails with an error. This allows fixing the embedding model config via admin panel. Previously server failed to start if embedding model was configured incorrectly. This prevented fixing the model config via admin panel. - Convert boolean string in config json to actual booleans when passed via admin panel as json before passing to model, query configs - Only create default model if no search model configured by admin. Return first created search model if its been configured by admin.	2024-10-30 14:00:21 -07:00
Debanjum	358a6ce95d	Defer turning cursor color to selected agents color for later Capability exists but idea needs to be investigated further	2024-10-30 14:00:21 -07:00
Debanjum	2ac840e3f2	Make cursor in chat input take on selected agent color	2024-10-30 14:00:21 -07:00
Debanjum	1448b8b3fc	Use 3rd person for user in research prompt to reduce person confusion Models were getting a bit confused about who is search for who's information. Using third person to explicitly call out on who's behalf these searches are running seems to perform better across models (gemini's, gpt etc.), even if the role of the message is user.	2024-10-30 13:49:48 -07:00
Debanjum	b8c6989677	Separate example from actual question in extract question prompt	2024-10-30 13:49:48 -07:00
Debanjum	86ffd7a7a2	Handle \n, dedupe json cleaning into single function for reusability Use placeholder for newline in json object values until json parsed and values extracted. This is useful when research mode models outputs multi-line codeblocks in queries etc.	2024-10-30 13:49:48 -07:00

... 2 3 4 5 6 ...

3915 commits