you-n-g
28ceb41e1c
fix: clear ws_ckp after extraction to reduce workspace object size ( #1137 )
2025-07-31 18:10:32 +08:00
you-n-g
7fc09169bc
feat: fallback to acceptable results ( #1129 )
...
* refactor: add is_acceptable, fallback logic and generify evolving agent
* refine lint
* small
* lint
* lint
* lint
* feat: add is_acceptable to CoSTEERMultiFeedback
* feat: add in-memory workspace checkpoint and recovery
* feat: preserve symbolic links in workspace checkpoints and recovery
* lint
* lint
* feat: limit workspace checkpoint to files under 100KB
* feat: add workspace checkpoint size limit setting
* prompt
* lint
2025-07-31 17:53:18 +08:00
Xu Yang
fd6cd3950c
fix: remove unused imports in data science scenario module ( #1136 )
2025-07-31 17:05:14 +08:00
Xu Yang
6a4998154d
feat: add time ratio limit for hyperparameter tuning in Kaggle settin… ( #1135 )
...
* feat: add time ratio limit for hyperparameter tuning in Kaggle settings and update evaluator logic
* add recommend time limit
* remove 25% hard line
* small fix
2025-07-31 16:47:19 +08:00
Xu Yang
305eff1c5e
feat: enhance timeout management and knowledge base handling in CoSTEER components ( #1130 )
...
* feat: enhance timeout management and knowledge base handling in CoSTEER components
* fix a little bug
* fix small bug
* fix a small bug
* Update rdagent/scenarios/data_science/loop.py
Co-authored-by: you-n-g <you-n-g@users.noreply.github.com >
* add scale check
* fix a small bug
* fix CI
* use dynamic chat_token_limit & remove repeated lines
* fix CI
* remove useless comment
* fix small bug
* update draft appendix
* fix prompt
* add code correctness as top priority
---------
Co-authored-by: you-n-g <you-n-g@users.noreply.github.com >
2025-07-31 13:06:28 +08:00
you-n-g
600d159e86
fix: update requirements.txt's streamlit ( #1133 )
2025-07-30 21:30:53 +08:00
XianBW
0ccb6e6d56
log time duration when call LLM and show timeout ratio of coding & running in trace UI ( #1128 )
2025-07-29 17:54:29 +08:00
Xu Yang
e0519b56ae
feat: add coder check and give more time ( #1127 )
...
* feat: introduce max_seconds_multiplier for timeout management across components
* avoid using assert in test
* hot fix a very small bug
* runner multiply twice
* revert feedback change
* add a switch to longer timeout
2025-07-29 14:59:38 +08:00
amstrongzyf
a6703aa313
feat: add hypo_critic and hypo_rewrite in proposal ( #1106 )
...
add hypo_critique and hypo_rewrite in proposal.py, controlled by `DS_RD_SETTING.enable_hypo_critique_rewrite` (default True)
2025-07-29 14:20:09 +08:00
Roland Minrui
eb70333b77
small fix to runner eval order to prevent both final decision and hyperparameter decision to be true
...
Co-authored-by: Xu <v-xuminrui@microsoft.com >
2025-07-29 13:31:35 +08:00
XianBW
4cee886c80
fix current_selection maybe empty ( #1125 )
2025-07-29 11:56:59 +08:00
XianBW
e19fe4b233
fix find last SOTA diff ( #1124 )
2025-07-29 10:56:41 +08:00
Linlang
bbcee389b4
chore: implement runtime_env func ( #1104 )
...
* implement runtime_env func for quant
* add runtime_info code
* add runtime env information to the prompt
* format with black
* optimize get_runtime_env code
* delete unnecessary files
* some refinement
* fix fin_quant bugs
---------
Co-authored-by: Xu Yang <peteryang@vip.qq.com >
2025-07-28 14:19:41 +08:00
Roland Minrui
514c191b75
fix: fix code diff bug ( #1115 )
...
* fix code diff bug
* fix minor bug
* reformat
* fix minor bug
* fix
* fix
* fix the prefix type
---------
Co-authored-by: Xu <v-xuminrui@microsoft.com >
2025-07-26 22:51:59 +08:00
XianBW
15fa03d415
fix sota_exp is None ( #1121 )
2025-07-25 11:53:26 +08:00
amstrongzyf
13f49062a4
fix a small bug in feedback ( #1120 )
2025-07-24 22:33:08 +08:00
XianBW
5025448e84
add lite curves figure ( #1119 )
2025-07-24 20:11:09 +08:00
XianBW
59a2d7f91d
fix curve ( #1118 )
2025-07-24 19:15:01 +08:00
XianBW
e92cd99450
chore: some ui features change ( #1117 )
...
* add some todos
* fix bug
* fix leaderboard download function
* summary stat fix
* show medal
* show SOTA Loop
* fix CI
2025-07-24 18:50:29 +08:00
Tim
7a5dd566f5
chore: avoid more traces during merge ( #1108 )
...
* avoid more traces during merge
* restore change
* remove duplicate experiments
* add dropped pool
* use scenario
* remove drop_pool
* use max_sota_retrieved_num
2025-07-24 18:27:23 +08:00
Xu Yang
5e959a5f59
feat: analyze feedback based on sota numbers ( #1116 )
...
* give sota numbers to feedback
* refine prompt
2025-07-24 14:17:52 +08:00
XianBW
bbb008c95e
fix cli command type checker ( #1114 )
2025-07-23 19:43:25 +08:00
Xu Yang
0c3d0cd327
fix: prompt yaml ( #1112 )
...
* small fix to prompt yaml
* shrink the content
* remove parallel.py
2025-07-23 15:03:59 +08:00
Linlang
067c5a3334
fix health check bug ( #1111 )
2025-07-23 12:29:09 +08:00
Xu Yang
a3c0f29451
feat: enable meta planner ( #1103 )
...
* enable meta planner
* fix a small bug
* ADD PLAN TO GEN
* remove ensemble in planner
* fix CI
* fix CI
* fix planner threshold
2025-07-23 12:20:06 +08:00
xuangu-fang
733a17dd84
feat: new prompt for auto-sota-selector ( #1109 )
...
* fix mismmatch typo
* new sota_select_submit prompt
* update
* polish prompt
2025-07-22 21:51:31 +08:00
Tim
9f7a4280dc
chore: show merge loop in figure ( #1105 )
2025-07-22 17:02:56 +08:00
Tim
1f142cc5d8
chore: update figure ( #1102 )
...
* chore: update figure
* add record id
2025-07-21 18:53:07 +08:00
Tim
68f961407e
chore: merge enable parallel ( #1093 )
...
* chore: merge enable parallel
2025-07-21 14:28:36 +08:00
XianBW
c326aa52b9
chore: add response format in chatbot ( #1099 )
...
* add response format in chatbot
* fix CI
2025-07-21 12:19:05 +08:00
you-n-g
810af369e0
feat: add loop ID mapping to trace nodes and update UI labels ( #1098 )
...
* Add loop index
* feat: add loop ID mapping to trace nodes and update UI labels
* lint
* doc lint
2025-07-21 11:02:37 +08:00
you-n-g
ab97dd8b77
feat: add extra_eval config and import_class for custom evaluators ( #1097 )
...
* feat: add extra_eval config and import_class for custom evaluators
* lint
* build: update litellm requirement to >=1.73 for get_valid_models
* refactor: remove *args/**kwargs from _create_embedding_inner_function signature
2025-07-20 16:31:43 +08:00
you-n-g
54d72e1dca
fix: use CoSTEERSettings for DSRunnerCoSTEERSettings ( #1096 )
...
* refactor: use CoSTEERSettings for DSRunnerCoSTEERSettings
* lint
2025-07-20 10:53:09 +08:00
Linlang
81d0ac7185
docs: update data science docs ( #1015 )
...
* update data science docs part1
* update data science docs part2
* update data science docs part3
* update data science docs part4
* update data science docs part5
* format with isort
* update data science docs part6
* add check environment scripts
* format with isort
* format with isort
* merge env_check to health_check
* format with isort
* format with black
* optimize code
* use boolean cli options
* replace fire with typer
* replace fire with typer
* optimizing parameter and variable naming
2025-07-19 18:35:24 +08:00
Tim
705eb07e81
chore: extend eda output ( #1086 )
2025-07-18 21:07:43 +08:00
amstrongzyf
b1b77bd98b
fix: add gpu_info in research phase ( #1094 )
2025-07-18 21:00:06 +08:00
Xu Yang
9b82bc10ad
fix: package and timer bug ( #1092 )
...
* fix package bug & timer bug
* fix typo
2025-07-18 16:22:57 +08:00
XianBW
0e5a9894c5
use enum value when generate hypothesis ( #1091 )
2025-07-18 16:22:34 +08:00
XianBW
3b57264d2c
fix ui bugs ( #1090 )
2025-07-18 16:01:39 +08:00
Xu Yang
51921c2344
fix: minor fix to runtime_environment ( #1089 )
...
* minor fix to funtime_environment
* add python runtime info header
2025-07-17 19:58:06 +08:00
XianBW
221fea1e15
chore: add a simple chatbot in LLM_LOG window in DataScience UI ( #1088 )
...
* refine time show
* updates
* add chatbot in llm_log_win
* use newest litellm
2025-07-17 19:29:25 +08:00
Xu Yang
77725913fa
feat: query & cache package_info ( #1083 )
...
* feat: add package query in draft.py (not yet enabled)
* feat: integrate package query into task_gen and cache runtime environment
- Remove pkg_query modifications from draft components
- Add package declaration requirement in task_gen prompts
- Add optional packages field to CodingSketch model
- Cache runtime_environment in scenario object for loop-wide reuse
- Parse packages from LLM response and generate runtime environment dynamically
* some refinement
* feat: merge default packages with CLI args in package_info.py
* fix: code style
---------
Co-authored-by: Qizheng Li <jenssenlee@163.com >
2025-07-17 18:35:36 +08:00
Tim
e389b4a396
chore: change default setting ( #1087 )
...
live_output = True
2025-07-17 17:19:10 +08:00
you-n-g
79d21589c1
fix: ignore class types when filtering workflow steps ( #1085 )
...
* fix: ignore class types when filtering workflow steps
* chore: add TODO for record method in RDLoop
* lint
2025-07-17 15:50:41 +08:00
amstrongzyf
e3679d6c0d
fix: refine the prompt to force complete code & refine the logic of running ( #1069 )
...
* change refine prompt for full code
* fix: fix the logic of running
* refine prompt
* fix some bugs
* fix
* add two guidelines
* refactor the code
* make costeer evaluator more logical
* refine eval prompt
* make costeer eval prompt markdown
* update code diff prompt
* correct pipeline
* feat: add apply_patch utility and update ret.py with patch functionality (#1071 )
* restore to the right version
* fix the docstring
* fix extract_output fcn
* add inplace parameter to apply patch
* remove enable_runner_iteration and make the eval prompt same as main
* refine runner eval prompt based on main
* Update rdagent/scenarios/data_science/dev/runner/prompts.yaml
* add wait_retry
* refactor: move enable_runner_code_diff to DSRunnerCoSTEERSettings as diff_mode
* reformat and remove enable_runner_code_diff
---------
Co-authored-by: yuanteli <1957922024@qq.com >
Co-authored-by: Xu <v-xuminrui@microsoft.com >
Co-authored-by: Jensen Lee <91518020+Jensen246@users.noreply.github.com >
Co-authored-by: you-n-g <you-n-g@users.noreply.github.com >
Co-authored-by: Qizheng Li <jenssenlee@163.com >
2025-07-17 11:59:27 +08:00
Shuo Yin
6469a0679a
docs: update docs/installation_and_configuration.rst for Configuration ( #1061 )
...
* Update README.md
* Update README.md
* Update README.md
* Update README.md
* Update README.md
* Update README.md
* Update README.md
* Update README.md
* Update README.md
* Update README.md
* Update README.md
* Update README.md
* Update README.md
* Update README.md
* Update README.md
* Update installation_and_configuration.rst
* Update installation_and_configuration.rst
* Update installation_and_configuration.rst
* Update docs/installation_and_configuration.rst
Co-authored-by: you-n-g <you-n-g@users.noreply.github.com >
* Update installation_and_configuration.rst
* Update installation_and_configuration.rst
---------
Co-authored-by: you-n-g <you-n-g@users.noreply.github.com >
2025-07-16 22:20:49 +08:00
Xu Yang
427f5d270e
fix: minor conflict in prompts ( #1081 )
2025-07-16 16:10:56 +08:00
Xu Yang
daea490941
fix: align scenario descriptions and include debug timeout ( #1079 )
...
* Align scenario descriptions and include debug timeout
- Updated config.py to support debug timeout configuration
- Synchronized prompts in exp_gen and scen modules
- Refactored proposal.py for consistency with new scenario descriptions
- Improved __init__.py for better scenario management
* remove running time in stdout
---------
Co-authored-by: Xu Yang <xuyang1@microsoft.com >
2025-07-16 14:30:17 +08:00
you-n-g
de34c99931
fix: refine prompt; runner focus on low hanging fruit ( #1076 )
...
* fix: refine prompts formatting and logic across data science scenarios
* Make runner evaluator more detailed
* feat: switch to system_debugger prompt and print early stopping stats
2025-07-16 12:29:57 +08:00
you-n-g
5496168d07
feat: add enable_cache toggle for UI data caching ( #1075 )
2025-07-15 22:26:06 +08:00