you-n-g
4f0b2be7a7
feat: init pydantic ai agent & context 7 mcp ( #1240 )
...
* feat: init pydantic ai agent & context 7 mcp
* feat: integrate MCP documentation search into data science pipeline evaluation
* fix: disable MCP documentation search and update related docstrings and defaults
* lint
* fix: correct prompt formatting and conditional blocks in pipeline_eval section
* lint
* feat: add query method to PAIAgent for synchronous agent execution
* fix: apply nest_asyncio for agent and update context7 query method
* lint
* lint
* lint
* lint
* docs: update MCP folder docstring and rename test class in test_pydantic.py
* refactor: centralize completion kwargs logic and update pydantic_ai integration
* fixbug
* typo
* fix: bug triggered by padantic-ai version backtracking.
---------
Co-authored-by: Linlang <Lv.Linlang@hotmail.com >
2025-09-13 10:25:02 +08:00
amstrongzyf
7f94e3a9c3
fix: fix chat_max_tokens calculation method to show true input_max_tokens ( #1241 )
...
* fix: use real max_input_length
* lint
* Update rdagent/oai/backend/litellm.py
Co-authored-by: you-n-g <you-n-g@users.noreply.github.com >
* fix lint
* lint
* lint
---------
Co-authored-by: you-n-g <you-n-g@users.noreply.github.com >
2025-09-12 15:09:20 +08:00
amstrongzyf
4de999c4a3
fix: revert 2 commits ( #1239 )
...
* revert commit:c51a3008f897416a7416f92d1286cbbf6a2291ae. PR 1179
* Revert "fix: change runner prompts (#1223 )"
This reverts commit db2ebc27d70663d6aa85e0b05d9b81f5f6443c17.
* fix ruff check error
* fix: change switch place for fix_seed_and_data_split
2025-09-11 16:14:48 +08:00
you-n-g
6c132204c1
fix: move task cancellation to finally block and fix subprocess kill typo ( #1234 )
2025-09-10 10:05:08 +08:00
you-n-g
262081b2a5
feat: add stdout into workspace for easier debugging ( #1236 )
...
* refactor: use get_truncated_stdout for consistent stdout handling across modules
* lint
* feat: add dump_stdout_type to DSRunnerCoSTEERSettings and use in eval
* fix: avoid circular import by moving DSRunnerEvaluator import inside method
2025-09-10 10:02:49 +08:00
Tim
33fde3a0bb
feat: offline selector ( #1231 )
...
* offline selector test
* fix score with tensor
* sort sota list
---------
Co-authored-by: Xu Yang <peteryang@vip.qq.com >
2025-09-09 11:56:42 +08:00
Star dust
337389bd3f
fix: add a switch for ensemble_time_upper_bound and fix some bug in main ( #1226 )
...
* change runner prompts
* v1
* ensemble_time_upper_bound
* lint
* fix inf bug in function prob_dis_torch() and ensemble prompts
* lint
* lint
---------
Co-authored-by: amstrongzyf <amstrongzyf@126.com >
2025-09-08 16:35:42 +08:00
you-n-g
2e09996154
fix: increase retry count in hypothesis_gen decorator to 10 ( #1230 )
2025-09-08 12:23:58 +08:00
you-n-g
39de6bd168
fix: handle ValueError in stdout shrinking and refactor shrink logic ( #1228 )
2025-09-07 23:24:45 +08:00
Tim
956df9906c
chore: make sure token size below limit ( #1225 )
...
* chore: make sure token size below limit
* refactor: refactor and add import
---------
Co-authored-by: amstrongzyf <amstrongzyf@126.com >
2025-09-04 15:07:42 +08:00
Star dust
2eaa69536c
fix: change runner prompts ( #1223 )
...
* change runner prompts
* v1
2025-09-04 11:31:42 +08:00
amstrongzyf
71689b2cac
fix: revert to v10 setting ( #1220 )
...
* feat: add runner_patience
* task gen prompts and torch version
* fix ensemble time bug
* small bug
* lint
* add torch
* fiix: remove useless code and change formula
* fix: rename parameter
---------
Co-authored-by: jingyuanlm <842442862@qq.com >
2025-09-03 16:37:33 +08:00
XianBW
4347c2e98c
fix: summary page bug ( #1219 )
...
* fix summary bug
* fix CI
2025-09-02 17:06:54 +08:00
you-n-g
dd4edb8331
feat: ui, support disable cache ( #1217 )
...
* fix: (drop) add cache_enable for functions in ui.utils
* fix: (drop) handle no logging, rename sota_exp to last_sota_exp for clarity in running_win
2025-09-01 15:58:05 +08:00
amstrongzyf
c8830f3f3d
fix: jinja problem of enumerate ( #1216 )
2025-09-01 14:50:26 +08:00
XianBW
88a1b08b8c
for ui, a simple fix ( #1215 )
2025-09-01 14:46:51 +08:00
you-n-g
f0b7fa2466
fix: allow prev_out keys to be None in workspace cleanup assertion ( #1214 )
...
* fix: allow prev_out keys to be None in workspace cleanup assertion
* fix: clean workspace only for non-None DSExperiment instances
2025-08-30 07:27:12 +08:00
amstrongzyf
19eb7ae138
fix: add missing self parameter to instance methods in DSProposalV2ExpGen ( #1213 )
2025-08-30 00:45:55 +08:00
you-n-g
ac31473bfc
fix: handle None output and conditional step dump in LoopBase execution ( #1212 )
2025-08-29 20:19:09 +08:00
you-n-g
fbc1beb927
fix: update fallback criterion ( #1210 )
...
* fix: update fallback criterion
* fix: ensure evo_fb is initialized and used correctly in fallback logic
* refactor: rename use_new_evo to should_use_new_evo for clarity
2025-08-29 17:07:41 +08:00
you-n-g
1b15b53c1a
feat: add option to enable hyperparameter tuning only in first eval loop ( #1211 )
...
* feat: add option to enable hyperparameter tuning only in first eval loop
* fix: use total_seconds() for accurate time calculations in evolution and tracking
2025-08-29 16:59:22 +08:00
XianBW
7735e121b4
chore: ui updates ( #1209 )
...
* show a workspace path string for easier copy, only percent existing columns when percent summary dataframe
* catch summary exception
* fix a bug
* fix ci
2025-08-29 15:05:15 +08:00
amstrongzyf
da8f61f874
fix: fix bug for hypo_select_with_llm when not support response_schema ( #1208 )
2025-08-28 14:36:23 +08:00
XianBW
11b0f1b9ee
fix ui bug ( #1207 )
2025-08-28 14:18:58 +08:00
you-n-g
64734fba4c
fix: move snapshot saving after step index update in loop execution ( #1206 )
2025-08-28 10:58:57 +08:00
amstrongzyf
40f26ac54d
feat: enable LLM‑based hypothesis selection with time‑aware prompt & colored logging ( #1122 )
...
* add hypo select by llm (without time)
* add time_info and log color
* select no smooth
* 2 hypo
* change select
* small change
* fix bug
* fix bug and add hypothesis router and begin flag
* fix bug v1
* fix bug v2
* fix feedback
* add new model
* add filter
* fix bug v3
* fix bug v4
* change prompts v2
* fix bug v5
* fix bug v6
* fix hypo
* fix some bug(sota socre, prompts, ensemble prompts ) and add path legth.
* fix bug v7
* fix bug v8
* fix bug v9
* fix bug v10
* reset to v10 and refine
* fix: translate to english
* fix bug
* fix: use differnet increase_stage for coder & runner
* fix: use different timeout_increase_stage for coder and runner.
* fix: revert logger
* fix: remove duplicate content
* feat: implement LLM-driven extra hypothesis selection and adjust logic
* refactor: relocate Hypothesis models below TraceChallenges
* remove torch fix bug
* fix small bug
* refactor: prefix internal methods with underscore for llm-based hypothesis selection
* add fix_seed_and_data_split and enable_simple_hypothesis
* lint
* refine proposal config order
* code review; comments
* lint
* merge int to float
---------
Co-authored-by: jingyuanlm <842442862@qq.com >
Co-authored-by: Young <afe.young@gmail.com >
2025-08-27 18:36:51 +08:00
XianBW
0018429dc4
chore: refine get_summary_df and sota exp stat in log utils ( #1201 )
...
* add best valid selector to get sota exp stat
* remove unused in summary & refine get_summary_df
* fix hypothesis table bug in trace
2025-08-26 18:31:21 +08:00
Linlang
b4e11f1895
style: convert CLI flags from snake_case to kebab-case ( #1198 )
...
Co-authored-by: Young <afe.young@gmail.com >
2025-08-25 20:07:01 +08:00
Xu Yang
3b57ea2666
fix: increase time default not controlled by LLM ( #1196 )
...
* fix: add longer timeout configuration for data science scenarios
* fix CI
2025-08-21 11:27:17 +08:00
XianBW
2f54199b44
fix: kaggle competition metric direction ( #1195 )
...
* fix leaderboard and competition direction
* ui fix
2025-08-20 16:40:53 +08:00
amstrongzyf
2173fc8c25
fix bug for embedding caused by PR 1188 ( #1194 )
2025-08-18 20:15:18 +08:00
Tim
08db63e0fe
chore: refine env command ( #1193 )
2025-08-18 18:16:16 +08:00
XianBW
4728df7f38
fix: ui bug ( #1192 )
...
* small bug
* fix ci
2025-08-18 17:48:49 +08:00
amstrongzyf
5785bcf2ba
fix: enable embedding truncation ( #1188 )
...
* fix: enable embedding auto truncation
* move embedding_utils to utils.embedding
* fix bug
* fix(embedding): improve text truncation with retry strategies and binary search optimization
* new way for embedding truncate
* fix ci errors
2025-08-18 17:20:41 +08:00
xuangu-fang
ceabfe9686
feat: enable to inject diversity cross async multi-trace ( #1173 )
...
* init multi_trace_async_ diversity inject
* lint
* add task in context for diversity injection, add abstract class for diversity injection
* lint
* feat: update diversity injection strategy and enhance sibling context handling in experiment generation
* add always inject
---------
Co-authored-by: Xu Yang <peteryang@vip.qq.com >
2025-08-18 12:43:21 +08:00
you-n-g
0885d4d218
change the trace number from 3 to 1 (simpler default value) ( #1191 )
2025-08-18 12:13:23 +08:00
XianBW
c3ef9e2d48
fix ui bugs ( #1190 )
2025-08-18 10:56:51 +08:00
you-n-g
00b588ffd4
fix: skip res_ratio check if timer or res_time is None ( #1189 )
2025-08-17 16:08:55 +08:00
Tim
0dae88e2cd
chore: add inference mode for model dump ( #1182 )
...
* chore: add inference mode for model dump
* check runtime environment
* update opened_trace_lines
---------
Co-authored-by: you-n-g <you-n-g@users.noreply.github.com >
2025-08-15 15:34:49 +08:00
you-n-g
bfbdbcfbbd
fix: insert await asyncio.sleep(0) to yield control in loop ( #1186 )
2025-08-15 13:27:29 +08:00
you-n-g
254eb56da3
refactor: use UI_SETTING.amlt_path for combined log folder path ( #1184 )
2025-08-14 21:51:05 +08:00
Tim
f87dc2a5b8
chore: catch log_workflow error ( #1183 )
2025-08-14 16:12:37 +08:00
amstrongzyf
5b785346f7
fix: refine prompts and add additional package info ( #1179 )
...
* refine prompts and add additional package info
* refine prompts to be specific for GBDT models
* minor refine prompts
* use include to replace duplicate info
* refine prompts
* refactor: import DSTrace from base and remove exp_gen __init__
* lint
---------
Co-authored-by: Young <afe.young@gmail.com >
2025-08-13 23:00:23 +08:00
Tim
aa2a5bae15
chore: add merge statistics ( #1176 )
...
* chore: add stat for merge
* use best selector
* use trace.hist
2025-08-13 22:59:34 +08:00
Chaos Yu
30d6afbb45
fix(graph): using assignment expression to avoid repeated function call ( #1174 )
...
* using assignment expression to avoid repeated function call
* format
2025-08-13 17:16:58 +08:00
Roland Minrui
b94414a8ed
feat: refine the logic of enabling hyperparameter tuning and add criteira ( #1175 )
...
* add 4 tuning criteria
* fix equation direction
* fix ci
* add comment
Co-authored-by: you-n-g <you-n-g@users.noreply.github.com >
---------
Co-authored-by: Xu <v-xuminrui@microsoft.com >
Co-authored-by: you-n-g <you-n-g@users.noreply.github.com >
2025-08-12 21:12:54 +08:00
you-n-g
137bdc15c5
fix: cancel tasks on resume and kill subprocesses on termination ( #1166 )
...
* fix: cancel tasks on resume and kill subprocesses on termination
* lint
* lint
2025-08-09 19:39:13 +08:00
XianBW
5b6a7d1d9b
fix exec_time in summary because of parallel ( #1171 )
2025-08-08 17:14:02 +08:00
XianBW
dbbf43558d
add auto sota selector llm log in trace ( #1169 )
2025-08-08 16:26:48 +08:00
Xu Yang
3fe2ba44df
feat: streamline hyperparameter tuning checks and update evaluation g… ( #1167 )
...
* feat: streamline hyperparameter tuning checks and update evaluation guidelines
* fix task_gen json check
2025-08-08 13:49:01 +08:00