Utsab Dahal
dc3941e331
fix: prevent JSON content from being added multiple times during retries ( #1255 )
2025-10-02 09:45:58 +08:00
Dex
8e126b0b57
fix(docs): update rdagent ui with correct params ( #1249 )
2025-09-23 16:20:00 +08:00
Tim
718a7ab779
fix: merge candidates ( #1254 )
...
* fix: add hypotheses_candidates
2025-09-22 21:04:07 +08:00
Xu Yang
4e6adc1f20
feat: update README with latest paper acceptance to NeurIPS 2025 ( #1252 )
...
* feat: update README with latest news and announcements
* update readme
2025-09-19 10:58:07 +08:00
Xu Yang
07c228ef9f
feat: add user interaction in data science scenario ( #1251 )
...
* feat: add interactor classes and user interaction handling for experiments
* update code
* use fragment retry mechanism instead of rerun()
* fix a bug
* integrate user instructions into proposal and coder
* fix CI
* fix CI
* feat: add approval option for user instructions submission
* feat: enhance user instructions handling in Task and DSExperiment classes
* fix CI
* add user instructions into hypothesis rewrite
* add interface to command line
---------
Co-authored-by: Bowen Xian <xianbowen@outlook.com >
2025-09-18 11:10:12 +08:00
Tim
121b2dbe48
chore: valid selector update ( #1248 )
...
* chore: add medal info
* return sota_exp_stat
* update hit check
* update experiment
* udpate experiment
* extract log
* remove old log folder
* update candidates
* keep highest score
* early stop if no medal candidate
2025-09-17 15:48:16 +08:00
you-n-g
e78398996c
fix: add json format response fallback to prompt templates ( #1246 )
...
* fix: add json format response fallback to prompt templates
* more json prompt
2025-09-15 23:27:24 +08:00
you-n-g
6577f61df0
fix: set requires_documentation_search to None to disable feature in eval ( #1245 )
2025-09-15 07:45:08 +08:00
you-n-g
4f0b2be7a7
feat: init pydantic ai agent & context 7 mcp ( #1240 )
...
* feat: init pydantic ai agent & context 7 mcp
* feat: integrate MCP documentation search into data science pipeline evaluation
* fix: disable MCP documentation search and update related docstrings and defaults
* lint
* fix: correct prompt formatting and conditional blocks in pipeline_eval section
* lint
* feat: add query method to PAIAgent for synchronous agent execution
* fix: apply nest_asyncio for agent and update context7 query method
* lint
* lint
* lint
* lint
* docs: update MCP folder docstring and rename test class in test_pydantic.py
* refactor: centralize completion kwargs logic and update pydantic_ai integration
* fixbug
* typo
* fix: bug triggered by padantic-ai version backtracking.
---------
Co-authored-by: Linlang <Lv.Linlang@hotmail.com >
2025-09-13 10:25:02 +08:00
amstrongzyf
7f94e3a9c3
fix: fix chat_max_tokens calculation method to show true input_max_tokens ( #1241 )
...
* fix: use real max_input_length
* lint
* Update rdagent/oai/backend/litellm.py
Co-authored-by: you-n-g <you-n-g@users.noreply.github.com >
* fix lint
* lint
* lint
---------
Co-authored-by: you-n-g <you-n-g@users.noreply.github.com >
2025-09-12 15:09:20 +08:00
amstrongzyf
4de999c4a3
fix: revert 2 commits ( #1239 )
...
* revert commit:c51a3008f897416a7416f92d1286cbbf6a2291ae. PR 1179
* Revert "fix: change runner prompts (#1223 )"
This reverts commit db2ebc27d70663d6aa85e0b05d9b81f5f6443c17.
* fix ruff check error
* fix: change switch place for fix_seed_and_data_split
2025-09-11 16:14:48 +08:00
you-n-g
6c132204c1
fix: move task cancellation to finally block and fix subprocess kill typo ( #1234 )
2025-09-10 10:05:08 +08:00
you-n-g
262081b2a5
feat: add stdout into workspace for easier debugging ( #1236 )
...
* refactor: use get_truncated_stdout for consistent stdout handling across modules
* lint
* feat: add dump_stdout_type to DSRunnerCoSTEERSettings and use in eval
* fix: avoid circular import by moving DSRunnerEvaluator import inside method
2025-09-10 10:02:49 +08:00
Tim
33fde3a0bb
feat: offline selector ( #1231 )
...
* offline selector test
* fix score with tensor
* sort sota list
---------
Co-authored-by: Xu Yang <peteryang@vip.qq.com >
2025-09-09 11:56:42 +08:00
Star dust
337389bd3f
fix: add a switch for ensemble_time_upper_bound and fix some bug in main ( #1226 )
...
* change runner prompts
* v1
* ensemble_time_upper_bound
* lint
* fix inf bug in function prob_dis_torch() and ensemble prompts
* lint
* lint
---------
Co-authored-by: amstrongzyf <amstrongzyf@126.com >
2025-09-08 16:35:42 +08:00
you-n-g
2e09996154
fix: increase retry count in hypothesis_gen decorator to 10 ( #1230 )
2025-09-08 12:23:58 +08:00
you-n-g
39de6bd168
fix: handle ValueError in stdout shrinking and refactor shrink logic ( #1228 )
2025-09-07 23:24:45 +08:00
Tim
956df9906c
chore: make sure token size below limit ( #1225 )
...
* chore: make sure token size below limit
* refactor: refactor and add import
---------
Co-authored-by: amstrongzyf <amstrongzyf@126.com >
2025-09-04 15:07:42 +08:00
Star dust
2eaa69536c
fix: change runner prompts ( #1223 )
...
* change runner prompts
* v1
2025-09-04 11:31:42 +08:00
amstrongzyf
71689b2cac
fix: revert to v10 setting ( #1220 )
...
* feat: add runner_patience
* task gen prompts and torch version
* fix ensemble time bug
* small bug
* lint
* add torch
* fiix: remove useless code and change formula
* fix: rename parameter
---------
Co-authored-by: jingyuanlm <842442862@qq.com >
2025-09-03 16:37:33 +08:00
XianBW
4347c2e98c
fix: summary page bug ( #1219 )
...
* fix summary bug
* fix CI
2025-09-02 17:06:54 +08:00
you-n-g
dd4edb8331
feat: ui, support disable cache ( #1217 )
...
* fix: (drop) add cache_enable for functions in ui.utils
* fix: (drop) handle no logging, rename sota_exp to last_sota_exp for clarity in running_win
2025-09-01 15:58:05 +08:00
amstrongzyf
c8830f3f3d
fix: jinja problem of enumerate ( #1216 )
2025-09-01 14:50:26 +08:00
XianBW
88a1b08b8c
for ui, a simple fix ( #1215 )
2025-09-01 14:46:51 +08:00
you-n-g
f0b7fa2466
fix: allow prev_out keys to be None in workspace cleanup assertion ( #1214 )
...
* fix: allow prev_out keys to be None in workspace cleanup assertion
* fix: clean workspace only for non-None DSExperiment instances
2025-08-30 07:27:12 +08:00
amstrongzyf
19eb7ae138
fix: add missing self parameter to instance methods in DSProposalV2ExpGen ( #1213 )
2025-08-30 00:45:55 +08:00
you-n-g
ac31473bfc
fix: handle None output and conditional step dump in LoopBase execution ( #1212 )
2025-08-29 20:19:09 +08:00
you-n-g
fbc1beb927
fix: update fallback criterion ( #1210 )
...
* fix: update fallback criterion
* fix: ensure evo_fb is initialized and used correctly in fallback logic
* refactor: rename use_new_evo to should_use_new_evo for clarity
2025-08-29 17:07:41 +08:00
you-n-g
1b15b53c1a
feat: add option to enable hyperparameter tuning only in first eval loop ( #1211 )
...
* feat: add option to enable hyperparameter tuning only in first eval loop
* fix: use total_seconds() for accurate time calculations in evolution and tracking
2025-08-29 16:59:22 +08:00
XianBW
7735e121b4
chore: ui updates ( #1209 )
...
* show a workspace path string for easier copy, only percent existing columns when percent summary dataframe
* catch summary exception
* fix a bug
* fix ci
2025-08-29 15:05:15 +08:00
amstrongzyf
da8f61f874
fix: fix bug for hypo_select_with_llm when not support response_schema ( #1208 )
2025-08-28 14:36:23 +08:00
XianBW
11b0f1b9ee
fix ui bug ( #1207 )
2025-08-28 14:18:58 +08:00
you-n-g
64734fba4c
fix: move snapshot saving after step index update in loop execution ( #1206 )
2025-08-28 10:58:57 +08:00
amstrongzyf
40f26ac54d
feat: enable LLM‑based hypothesis selection with time‑aware prompt & colored logging ( #1122 )
...
* add hypo select by llm (without time)
* add time_info and log color
* select no smooth
* 2 hypo
* change select
* small change
* fix bug
* fix bug and add hypothesis router and begin flag
* fix bug v1
* fix bug v2
* fix feedback
* add new model
* add filter
* fix bug v3
* fix bug v4
* change prompts v2
* fix bug v5
* fix bug v6
* fix hypo
* fix some bug(sota socre, prompts, ensemble prompts ) and add path legth.
* fix bug v7
* fix bug v8
* fix bug v9
* fix bug v10
* reset to v10 and refine
* fix: translate to english
* fix bug
* fix: use differnet increase_stage for coder & runner
* fix: use different timeout_increase_stage for coder and runner.
* fix: revert logger
* fix: remove duplicate content
* feat: implement LLM-driven extra hypothesis selection and adjust logic
* refactor: relocate Hypothesis models below TraceChallenges
* remove torch fix bug
* fix small bug
* refactor: prefix internal methods with underscore for llm-based hypothesis selection
* add fix_seed_and_data_split and enable_simple_hypothesis
* lint
* refine proposal config order
* code review; comments
* lint
* merge int to float
---------
Co-authored-by: jingyuanlm <842442862@qq.com >
Co-authored-by: Young <afe.young@gmail.com >
2025-08-27 18:36:51 +08:00
XianBW
0018429dc4
chore: refine get_summary_df and sota exp stat in log utils ( #1201 )
...
* add best valid selector to get sota exp stat
* remove unused in summary & refine get_summary_df
* fix hypothesis table bug in trace
2025-08-26 18:31:21 +08:00
Linlang
b4e11f1895
style: convert CLI flags from snake_case to kebab-case ( #1198 )
...
Co-authored-by: Young <afe.young@gmail.com >
2025-08-25 20:07:01 +08:00
Xu Yang
3b57ea2666
fix: increase time default not controlled by LLM ( #1196 )
...
* fix: add longer timeout configuration for data science scenarios
* fix CI
2025-08-21 11:27:17 +08:00
XianBW
2f54199b44
fix: kaggle competition metric direction ( #1195 )
...
* fix leaderboard and competition direction
* ui fix
2025-08-20 16:40:53 +08:00
amstrongzyf
2173fc8c25
fix bug for embedding caused by PR 1188 ( #1194 )
2025-08-18 20:15:18 +08:00
Tim
08db63e0fe
chore: refine env command ( #1193 )
2025-08-18 18:16:16 +08:00
XianBW
4728df7f38
fix: ui bug ( #1192 )
...
* small bug
* fix ci
2025-08-18 17:48:49 +08:00
amstrongzyf
5785bcf2ba
fix: enable embedding truncation ( #1188 )
...
* fix: enable embedding auto truncation
* move embedding_utils to utils.embedding
* fix bug
* fix(embedding): improve text truncation with retry strategies and binary search optimization
* new way for embedding truncate
* fix ci errors
2025-08-18 17:20:41 +08:00
xuangu-fang
ceabfe9686
feat: enable to inject diversity cross async multi-trace ( #1173 )
...
* init multi_trace_async_ diversity inject
* lint
* add task in context for diversity injection, add abstract class for diversity injection
* lint
* feat: update diversity injection strategy and enhance sibling context handling in experiment generation
* add always inject
---------
Co-authored-by: Xu Yang <peteryang@vip.qq.com >
2025-08-18 12:43:21 +08:00
you-n-g
0885d4d218
change the trace number from 3 to 1 (simpler default value) ( #1191 )
2025-08-18 12:13:23 +08:00
XianBW
c3ef9e2d48
fix ui bugs ( #1190 )
2025-08-18 10:56:51 +08:00
you-n-g
00b588ffd4
fix: skip res_ratio check if timer or res_time is None ( #1189 )
2025-08-17 16:08:55 +08:00
Tim
0dae88e2cd
chore: add inference mode for model dump ( #1182 )
...
* chore: add inference mode for model dump
* check runtime environment
* update opened_trace_lines
---------
Co-authored-by: you-n-g <you-n-g@users.noreply.github.com >
2025-08-15 15:34:49 +08:00
you-n-g
bfbdbcfbbd
fix: insert await asyncio.sleep(0) to yield control in loop ( #1186 )
2025-08-15 13:27:29 +08:00
you-n-g
254eb56da3
refactor: use UI_SETTING.amlt_path for combined log folder path ( #1184 )
2025-08-14 21:51:05 +08:00
Tim
f87dc2a5b8
chore: catch log_workflow error ( #1183 )
2025-08-14 16:12:37 +08:00