ReZero-Search-LLM-Agent-Fork

Commit Graph

Author	SHA1	Message	Date
thinhlpg	c153652856	docs: update README with new image	7 months ago
thinhlpg	ad18169d77	docs: enhance README with demo GIF	7 months ago
thinhlpg	89e07bc02d	chore: chore: remove unused code and dependencies	7 months ago
thinhlpg	5eabd121a3	docs: update README	7 months ago
thinhlpg	0b4bf54833	feat: update demo from DeepSearch to ReZero, adjusting related logging and UI components	7 months ago
thinhlpg	9738b80353	feat: update max generations and output length in evaluation scripts, add memory fraction to server launch	7 months ago
thinhlpg	7ee65269fb	feat: add new evaluation notebook for model testing and checkpoint evaluation	7 months ago
thinhlpg	bec864038b	feat: increase max tokens and new tokens in evaluation scripts	7 months ago
thinhlpg	dfa420fa49	feat: expand Makefile with serving and evaluation commands	7 months ago
thinhlpg	6ba963aca3	feat: streamline data preparation in Makefile with a single command	7 months ago
thinhlpg	424459d840	feat: update evaluation scripts to enhance model configuration and dataset loading, including increased max tokens and added logging	7 months ago
thinhlpg	bf9f2c4102	docs: update README with setup instructions, quick demo, and data preparation steps for better clarity and usability	7 months ago
thinhlpg	d7cdb6c917	chore: remove unused scripts	7 months ago
thinhlpg	1e7514f98e	chore: remove outdated documentation files to clean up project structure	7 months ago
thinhlpg	333d1e596e	feat: add prepare-dev-data target and script for Musique dev data transformation	7 months ago
thinhlpg	504f0c6c8e	feat: update reward_em_chunk to match only the LAST required paragraph of the reasoning chain and adjust related tests	8 months ago
thinhlpg	358875a035	feat: enhance reward_em_chunk function to match multiple paragraphs, add test	8 months ago
thinhlpg	2df9f39fda	feat: update model configuration (longer context) and dataset loading logic for improved performance and flexibility	8 months ago
thinhlpg	4a1d45271d	feat: add scripts for musique data processing	8 months ago
thinhlpg	74aa673866	chores: add cook notebook for musique and model reasoning pattern	8 months ago
thinhlpg	14ef79a4f5	feat: [WIP] add bench scripts	8 months ago
thinhlpg	bd02305efb	chores: add cook notebooks	8 months ago
thinhlpg	d8e949ec7c	feat: add Tavily search tab and integrate TavilyClient for web search functionality	8 months ago
thinhlpg	41b7889a30	feat: integrate QA dataset loading and display gold answers in Gradio interface	8 months ago
thinhlpg	7376f596a5	feat: add Gradio demo for DeepSearch and update configuration settings	8 months ago
thinhlpg	7ff3623102	chore: update .gitignore, modify Makefile for installation, and add pyproject.toml for project configuration	8 months ago
thinhlpg	eebf914a81	refactor: moved modules from src/deepsearch to src/	8 months ago
thinhlpg	0f662d4330	refactor: moved FlashRAG submodule from src/ to third_party/	8 months ago
thinhlpg	55f34b8503	feat: add FlashRAG as submodule	8 months ago
thinhlpg	2fec4f2f42	refactor: change repo stucture (move code from src/ to src/deepsearch)	8 months ago
thinhlpg	e3163081a0	docs: add experiment log for llama-3.2-3b-instruct experiments	8 months ago
thinhlpg	010957cd99	feat: disable randomization option to get_qa_dataset function by default	8 months ago
automaticcat	56911a73f9	Update README.md	8 months ago
thinhlpg	1a18cd7bfd	feat: update training and evaluation configurations (editable agent generation scripts) Increased max_generations parameter in agentic_generate and run_eval functions for improved output flexibility.	8 months ago
thinhlpg	77f121662f	test: add tests for reward_retry function scenarios	8 months ago
thinhlpg	c8714e0f6b	feat: enhance reward_retry function to handle missing answer tags Added logic to return 0 if the final message from the assistant does not contain answer tags (no matter how hard you try, you won't get anything if no result 💀)	8 months ago
thinhlpg	bf480574a2	fix: minor bug	8 months ago
thinhlpg	3081d6e36b	test: added tests for new reward functions: search strategy and search diversity	8 months ago
thinhlpg	4de31e0f30	feat: expand reward functions with new strategies and diversity checks - Added reward functions for search strategy and search diversity - Updated reward_format to include validation for proper message endings.	8 months ago
thinhlpg	d0e6068055	fix: strengthen reward correctness logic to handle final message is not asnwer form assistant. Also update logs for reward functions for better debug - Added 'logs/' directory to .gitignore to exclude log files. - Introduced log_chat_state function to log chat states and rewards to JSONL files. - Updated reward functions to log chat states with validation results for better tracking and debugging.	8 months ago
thinhlpg	1bd609dfae	test: enhance reward correctness tests with validation logic - Updated test cases to include role and tag validation for assistant messages. - Ensured that only properly formatted messages with answer tags are accepted. - Added new test for validating various incorrect formats and their expected outcomes.	8 months ago
thinhlpg	338655e563	feat: refine user prompt logic for improved clarity and structure	8 months ago
thinhlpg	6d994feeb2	feat: enhance evaluation scripts for base and LoRA models	8 months ago
thinhlpg	da60b52bd1	feat: refactor download and upload scripts for improved argument handling (more notebook friendly :D)	8 months ago
thinhlpg	fa3c0562fe	feat: add evaluation scripts for base and LoRA models - Introduced `eval_base.py` for evaluating base model performance. - Introduced `eval_lora.py` for evaluating LoRA model performance with additional LoRA weight handling.	8 months ago
thinhlpg	1047e2fa1c	chore: update .gitignore and requirements for unsloth versions	8 months ago
thinhlpg	83f86869f6	chore: update .gitignore and add new toys data files	8 months ago
thinhlpg	133cb1ab90	test: add Qwen tokenizer adapter tests Implemented unit tests for the Qwen tokenizer adapter, including format handling, mask generation, and multi-turn conversation support	8 months ago
thinhlpg	6efe01d5ff	chore: update Makefile and requirements for testing - Added a 'test' target in Makefile to run unit tests using pytest. - Included 'wandb' in requirements.txt for experiment tracking.	8 months ago
thinhlpg	af7f38c792	feat: add code for qwen architecture	8 months ago

1 2

78 Commits (c153652856ca428584c7f6f54c19702c0ca28ab1) All Branches Search

78 Commits (c153652856ca428584c7f6f54c19702c0ca28ab1)

All Branches