Xu Yang
5fecefb802
make timeout fail a hyper parameter ( #859 )
...
Co-authored-by: Xu Yang <xuyang1@microsoft.com >
2025-05-08 21:30:44 +08:00
XianBW
6fbe213f3d
change method of cut log trace ( #858 )
2025-05-08 21:07:33 +08:00
Roland Minrui
24c4eb7225
fix: refine the time/memory constraints prompt in hypothesis proposal ( #856 )
...
* refine prompt
* refine the wording
* add ratelimit retry to align with the suggested wait seconds
* add max retry to 0
* don't delete hist
---------
Co-authored-by: Xu <v-xuminrui@microsoft.com >
Co-authored-by: Xu Yang <xuyang1@microsoft.com >
2025-05-08 21:05:07 +08:00
Linlang
eed4331f0f
fix: fixed CI execution failures caused by document builds ( #857 )
...
* fixed CI execution failures caused by document builds
* add comments
2025-05-08 20:35:13 +08:00
XianBW
9df59eb7f0
add cache for summary, remove workspace dependency when generate summary ( #854 )
2025-05-08 12:24:48 +08:00
you-n-g
e1d08cfe73
fix: adjust ds_trace lookup and add stderr redirect to mlebench command ( #853 )
...
* fix: adjust ds_trace lookup and add stderr redirect to mlebench command
* style: reformat SOTA experiment lookup in ds_trace.py
* feat: add DS_RD_SETTING pipeline to MergeExpGen success message
2025-05-08 02:40:49 +08:00
Xu Yang
45b260ba37
fix: trace list but ( #852 )
2025-05-07 21:28:24 +08:00
XianBW
8ee4c72bab
add aide.py ( #851 )
2025-05-07 19:59:40 +08:00
Xu Yang
668ac16150
feat: revert draft stage into a soft decay in hypothesis selection ( #849 )
...
* revert drafting
* update hypothesis rank logic
* prioritize time constraint in task design
* refine trace_desc and feedback problem prompt
* refine experiment_and_feedback_list_after_init
* fix DSHypothesis default parameter and print logic
* refine the selection weight
* merge simple_trace and trace
* refine weight and prompt
* refine sample logic
* fix CI
* robust code
---------
Co-authored-by: WinstonLiyte <1957922024@qq.com >
Co-authored-by: Xu <v-xuminrui@microsoft.com >
Co-authored-by: Xu Yang <xuyang1@microsoft.com >
2025-05-07 19:05:50 +08:00
Tim
3b7bafae12
chore: dump test data ( #850 )
...
* debug path for model dump
* update prompt
2025-05-07 17:24:50 +08:00
you-n-g
cda656b4ca
fix: non-exist variable test_eval.py ( #847 )
2025-05-07 04:36:00 +08:00
you-n-g
62e5953a6a
fix: wrong variable test_eval.py ( #846 )
2025-05-07 04:17:57 +08:00
XianBW
0047114a09
fix bug ( #845 )
2025-05-06 18:15:20 +08:00
Tim
46dec7b624
feat: custom data ( #810 )
...
* custom data
* fix: simplify competition check and log local description file
* no sample data
* feat: add test evaluation module with error handling support
* fix: update eval path to use eval_sub_dir and add valid_check TODO
* refactor: add MLETestEval check to conditionally run grading steps
* avoid blank stdout
* valid in testeval
* rename test.csv to avoid conflict
* Support Disabling sample submission
* refactoring
* fix: remove DS_KAGGLE_DATA and update prompt instructions
* add try for grade
* ignore submission
* fix: remove tee from eval command and warn about pipeline exit code detection
* optional to use raw description
* support old data
* add execution result to stdout
* add metric to raw description
* custom data explain
* add debug_path
* rst update
---------
Co-authored-by: Young <afe.young@gmail.com >
2025-05-06 16:00:13 +08:00
XianBW
eda4a2a807
chore: ui change ( #844 )
...
* ui change
* ui change
2025-05-06 11:34:21 +08:00
you-n-g
e8c39eb63d
feat: refine merge ( #842 )
...
* feat: add search_type param and customize success trial desc
* docs: simplify trial descriptions in share.yaml
* lint
2025-05-04 23:17:58 +08:00
you-n-g
fd98d29982
fix: adapting UI to mock trace ( #841 )
...
* fix: return first index if 'SOTA Exp Score (valid)' is empty
* feat: add get_state_data_range helper for loops and slider bounds
* fix: exclude batch embedding tag from log filtering
* feat: add feedback support in sota_experiment and adjust merge flow
* fix: use fb function for merging experiments
* style: remove extra whitespace and reformat code
2025-05-04 15:51:32 +08:00
Xu Yang
5cba25505e
fix new draft bugs ( #840 )
2025-04-30 18:53:27 +08:00
Haoran Pan
acc15caf7b
feat: reanalyze competition info & pipeline coding evaluator prompt ( #837 )
...
* update coding evaluator prompt similar to feedback
* reanaylyzing competition description when three sonsecutive coding failures
* update reanalyzing competition implementation
* fix bug
* update prompts and reanalyze
* fix bugs
* ci issue
* improve some code
* fix CI
---------
Co-authored-by: Xu Yang <peteryang@vip.qq.com >
Co-authored-by: Xu Yang <xuyang1@microsoft.com >
2025-04-30 17:36:35 +08:00
Roland Minrui
5f3da39331
feat: add drafting pipeline ( #832 )
...
* init commit
* add drafting prompt
* complete the drafting
* remove scenario problems from proposal
* rename prompts_drafting.yaml
* fix bug
* fix DSHypothesis print bug
* add failed drafting exp to prompt
* fix small bug
* use get_task_information() for task design
* resolve all comments
* add problem_desc to pesudo hypothesis
---------
Co-authored-by: Xu <v-xuminrui@microsoft.com >
2025-04-30 16:25:32 +08:00
you-n-g
ec51bb94b6
feat: trace merging ( #836 )
...
* feat: runnalbe -- add exp_gen_cls param, get_leaves and merge exp gen functionalities
* fix: remove unused scenario_desc and update YAML task labels
* feat: override selection and update merge task description
* lint
* lint
* lint
* lint
* lint
* fix: log competition setting to enable mle_summary
* fix name error
2025-04-29 09:30:45 +08:00
Xu Yang
7cbf5bf56c
when restart with kb and different pkl path, we need to initialize the kb from json file ( #835 )
2025-04-28 15:51:51 +08:00
Tim
6be059dd5d
chore: log cost object ( #829 )
2025-04-25 19:26:08 +08:00
Xu Yang
686437d671
add pipeline to hypothesis spec ( #828 )
2025-04-25 18:29:14 +08:00
Roland Minrui
50d3b7ce61
feat: propose hypothesis across multiple parts in pipeline ( #827 )
...
* refine hypothesis for pipeline
* refine component selection prompt
* update feedback prompt
* remove feat_eng in pipeline coding
---------
Co-authored-by: Xu <v-xuminrui@microsoft.com >
Co-authored-by: Xu Yang <peteryang@vip.qq.com >
2025-04-25 17:17:53 +08:00
Xu Yang
24c4c6755e
fix: add time to timer when api timeout bug ( #826 )
...
* If not use the session stored timer, need to replace it with the default timer
* when api timeout, add the waiting time to timer
2025-04-25 10:50:27 +08:00
Xu Yang
2b821884e7
If not use the session stored timer, need to replace it with the default timer ( #825 )
2025-04-25 00:39:31 +08:00
Yuante Li
7f90e44798
fix: fix a bug in docker result extraction ( #824 )
...
* fix a bug in docker result extraction
2025-04-24 19:02:38 +08:00
XianBW
f8af5c0682
can set reasoning_effor=None in chat_model_map ( #823 )
2025-04-24 18:47:24 +08:00
XianBW
f909b1b6bc
feat: using different chat model in different part ( #822 )
...
* using model in the chat_model_map in one tag
* add replace timer to DS loop
* fix CI
* fix CI
* add more custom config in chat_model_map
* fix CI
* fix CI
* fix CI
---------
Co-authored-by: Xu Yang <xuyang1@microsoft.com >
2025-04-24 18:22:51 +08:00
Yuante Li
44ccee864d
fix: fix model input shape bug and costeer_model bug ( #821 )
...
* fix model input shape bug and costeer_model bug
* fix a bug
2025-04-24 12:48:25 +08:00
Xu Yang
0ce1ca74a0
fix: update runner max loop to 1 in DS scenario ( #820 )
2025-04-23 19:15:29 +08:00
Yuante Li
18a115b4e0
fix: fix some minor bugs in qlib scenario ( #817 )
...
* fix some bugs
* fix a bug
* fix a bug in qlib frontend
* fix ci
* fic ci
* fix qlib Dockerfile
2025-04-23 18:48:02 +08:00
Haoran Pan
364e3c9f61
update qrcode ( #818 )
2025-04-23 16:12:48 +08:00
Yuante Li
39e25d28a1
fix: improve eval alignment check (e.g. small-scale finetuning) ( #802 )
...
* fix
* fix
2025-04-22 18:47:24 +08:00
Xu Yang
0f2399667b
feat: add mlflow logger in RD loop to log ( #815 )
...
* add mlflow logger in DS loop
* fix CI
* fix CI
---------
Co-authored-by: Xu Yang <xuyang1@microsoft.com >
2025-04-22 01:23:40 +08:00
Xu Yang
29b1cbc316
feat: archive python and csv files in workspace to maintain results ( #814 )
...
* archive workspace also
* remove non python csv and md files in workspace to avoid big workspace dump
* FIX ci
2025-04-21 16:34:29 +08:00
XianBW
1c9b00474b
fix: align competion_full_desc and scenario_all_desc, remove redundant info in problems proposal ( #808 )
...
* align competition desc & scenario desc string
* remove competition_desc when having used scenario_desc in problem gen
* fix bug
* remove redundant competition desc in naive expgen
* improve proposal prompt
* modify phrase
---------
Co-authored-by: Xu Yang <peteryang@vip.qq.com >
2025-04-18 16:48:26 +08:00
XianBW
6f3e1a5b8a
add select lite and select best button ( #809 )
2025-04-18 16:28:31 +08:00
Xu Yang
2b7122f41f
fix: bug fix in timer start ( #807 )
2025-04-18 16:16:28 +08:00
Xu Yang
6646613359
fix: bug in problem identification ( #806 )
2025-04-18 15:46:11 +08:00
Roland Minrui
6fe9be19cd
feat: idea pool integrated to exp_gen & add timer to RD-Agent & pause-resume to RD-loops ( #795 )
...
* update all code
* update all code
* dump knowledge base
* rename the tag
* add timer to RD-Agent
* fix CI
* fix CI
* use batch embedding
* fix a small bug
* fix prompt bug
* feat: add pause resume to handle K8S cluster pause (#804 )
* add resume to cluster running
* fix non-pickle problem
* fix a small bug
* fix a small bug
* avoid shutil move error
* refine the logic
* move knowledge base out of session
* avoid mistake information to pipeline coding
* avoid load and dump in steps
* archive the right folder
* small improvement
* avoid restart when timer is already started
* fix CI
---------
Co-authored-by: Xu Yang <xuyang1@microsoft.com >
---------
Co-authored-by: Xu Yang <peteryang@vip.qq.com >
Co-authored-by: Xu Yang <xuyang1@microsoft.com >
Co-authored-by: Xu <v-xuminrui@microsoft.com >
2025-04-18 14:01:03 +08:00
Linlang
6d56061341
chore: modify kaggle docs & Adding ds_loop at program entry ( #786 )
...
* modify kaggle docs
* optimise code based on comments
* Update docs/scens/kaggle_agent.rst
* fix docs build error
---------
Co-authored-by: you-n-g <you-n-g@users.noreply.github.com >
2025-04-17 22:02:49 +08:00
you-n-g
d46d27375d
refactor: use remove_eda_part for EDA cleanup, fix diff eval ( #800 )
2025-04-17 16:19:26 +08:00
Linlang
353d8f05ef
update wechat qrcode ( #798 )
2025-04-16 21:39:56 +08:00
XianBW
8bd48c3dde
fix retry when hypothesis gen ( #796 )
2025-04-16 19:01:46 +08:00
you-n-g
a974dd9ba3
refactor: use dynamic input path and update template loader ( #792 )
...
* refactor: use dynamic input path and update template loader
* fix: update include syntax for data source in prompts.yaml
* add customization path
* docs: update prompts for ensemble scoring and metric direction
* chore: remove obsolete data_science/share.yaml file
2025-04-16 18:11:46 +08:00
Xu Yang
829f5534ca
feat: raise error when timeout in api call ( #793 )
...
* small change to try catch in backend
* fix CI
* add timeout tolerance to 3
2025-04-16 13:50:20 +08:00
you-n-g
e05a645ad4
feat: refine prompt ( #760 )
...
* style: Simplify language and improve clarity in prompts and share.yaml
* style: Update prompt wording for clarity in raw_data_loader
* style: Simplify conditional logic in task_gen system prompt
* refactor: Update prompts and proposal for component output format handling
* fix: Correct grammar and add clarification in prompts.yaml
* feat: Include coding guidelines in data science component prompts
* lint
2025-04-16 09:35:30 +08:00
you-n-g
b18ed7ce27
fix: update metric direction to return bool ( #791 )
2025-04-15 17:10:31 +08:00