mirror of
https://github.com/NicolasBohn/NexQuant.git
synced 2026-08-01 17:37:43 +00:00
7f4c2d18c6
* refine ds modal for more cases: eval and es * update model template * prompts for model and ensemble * fix a bug * fix a bug * init: ds workflow evovingstrategy * Adding ensemble (#505) * Initial Draft * Updating logic for init * Revising * Successful Testing * Updating to use the latest & right class * bug: bug-fixing for testing * data science loop changes * data science loop base * ds loop feedback * fix * remove measure_time because it's duplicated (in LoopBase) * add the knowledge query for data_loader & feature * edit ds workflow evaluator * data_loader bug fix * stop evolving when all tasks completed * llm app change * fix break all complete strategy * Adding queried knowledge (#508) Co-authored-by: XianBW <36835909+XianBW@users.noreply.github.com> * fix loop bug * ds workflow evaluator; test; refine prompts * workflow spec * fix ci * feature task changes * ds loop change * fix a bug in feat * add query knowledge for model and workflow * llm_debug info(for show) using pickle instead of json * remove NextLoopException * loop change * coder raise CoderError when all sub_tasks failed * rename code_dict to file_dict in FBWorkspace * add CoSTEER unittest * now show self.version in Task.get_task_information(), simplify CoSTEER sub tasks definition * remove some properties in ModelTask, add model_type in it. * fix llm app bug * llm web app bug fix * ds loop bug fix * fix: give component code to feature&ens eval * loop catch error bug * rename load_from_raw_data to load_data * feat: Add debug data creation functionality for data science scenarios * support local folder (#511) * support local folder * remove unnecessary random * KaggleScen Subclass * small fix * use template for style description * update default scen to kaggle * update sample data script * make sure frac < 1 * fix a bug * feature spec changes * fix * changeimport order * clear unnecessary std outputs * fix a typo * create sample folder after unzip kaggle data * feature/model test script update * Align the data types across modules. * fix a bug in model eval * show line number * move sample entry point to app * spec & model prompt changes * Refine the competition specification to address the data type problem and the coherence issue. * fix some bugs * add file filter in FBworkspace.code property * support non-binary prediction * avoid too much warnings * fix a bug in ensemble module * filtered the knowledge query in all modules * delete RAG in idea proposal * refine the code in ensemble * show exp workspace in llm_st * exp_gen bug fix * feedback bug fix * use `feature` instead of `feat01` * Trace & method of judging if exp is completed change * fix a bug in package calling and execute ci * fix code * bug fix * bug fix * fix a bug * fix some bugs * fix a bug * refactor: Enhance error handling and feedback in data science loop * support different use_azure on chat and embedding models * multi-model proposal logic * fix a small syntax error * loopBase and some changes * ensemble scores change * fbworkspace.code -> .all_codes * use all model codes in workflow coder * check scores.csv's keys(model_names) * model name changes * add a todo in ensemble test * sota_exp changes * give model info in exp gen * add runner time limit * config using debug data or not in evals * exp to feedback base * add feature code when writing model task * small problem * copying during sampling * update * refactor: Simplify code handling and improve workspace management * model part output fix * print model's execution time * bug fix * ensemble test fix * ens small change * ens_test bug fix * Refine partial expansion logic to display only a few subfolders when their structure is uniform, improving readability in nested directories. * several update on prompts * sample subfolders * Filter the stdout after code execution to remove irrelevant information e.g. progress bars, whitespace characters, excessive line breaks. * Add some more prompts and comments * several update on the first init rounds * model timeout as error * fix pattern of getting model codes in workspace * small bux fix on model prompts * remove get_code_with_key since we have regex pattern * fix: Correct tqdm progress bar update logic in LoopBase class * feat: Add diff generation and enhance feedback mechanism in data science loop * update some fix to model and workflow prompts * refine the logic of progress bar filter * add last_successful_exp in exp_gen * fix a one line bug * add a hint in prompt * fix data sample for bms * fix data sample for bms * hypothesis small fix * crawler readme update * fix component gen * fix bug * annotation change * load description.md if it exists * refactor: Simplify SOTA description handling in feedback and prompts * refactor: Use shared templates for feedback and experiment descriptions * change webapp for model codes changes * update proposal * add timeout message for docker run output * fix * refine the code in docker time processing * use .shape instead of len() when do shape eval * won't change size during iteration * support bson sample * sample support jsonl and bson * add former_code to coder prompts * a little speed us in debug data creating * filter progress bar when eval ens and main * avoid costeer makes no change to former code * fix several log error * add timeout judge threshold * fix some bugs in the evaluation of component output shapes * File structure for supporting litellm (#517) Co-authored-by: Young <afe.young@gmail.com> * ignore submission and show processing * ignore submission and show processing * add efficiency notice * refactor: Enhance error message with detailed feedback summary * refactor: Simplify component handling in DSExpGen class * refactor: Update code structure and add docstring for clarity * reserve one sample to each label in data sampling * add Evaluation info * refine costeer code to avoid giving same code twice * use raw_description as plain text * add a prompt hint to avoid same dict key * model task name bug in first model exp gen * fix a typo * add some debug info in costeer tests * task init change * enhance data sampling * refine the code in data_loader * more reasonable loop * fix a bug in data folder description * add error msg & traceback to execution feedback * fix llm error msg detection * add task information to costeer eval & add cache to docker run(use zipfile to store the whole workspace) * fix CI first round * fix CI second round * use txt to store test script to avoid pytest * remove zipfile in requirements * add azure.identity to requirements * ignore debug web page * component test changes * remove redundent task_desc in model coder * feat: Add APE module and prompts for automated prompt engineering * fix: Update .gitignore and improve text formatting in eval.py * refactor: Update print output and improve code comments and imports * style: Fix string formatting and import order in ape.py and fmt.py * exclude ape * add a data folder notice * reduce unnecessary output to stdout * refine the code of describe_data_folder * fix ci * style: streamlit style update (#522) * streamlit style update * fix import * fix format * fix llm_st loop progress bar * debugapp small change * fix model str * refine some prompts * fix model str * fix CI * refine the logic associated with the data_folder * fix ci * small change * set filter_progress_bar as default in execute * model proposal with workflow * add submission check in workflow eval * fix bug * small change * fix CI * fix CI * refactor: Move generate_diff to utils and update DSExpGen logic * more reasonable prompt describing metric direction * fix a minor jinja2 bug * quick fix exp_gen bugs * fix the following bug * fix * fix some bugs * remove workflow from model * add pending_tasks_list in data science to enable coding model and workflow * refine the code for handling JSON-formatted data descriptions * assert with information * ensure correct csv file name * add logging to help record the output * log competition * add log tag for debug llm app * test: Test ds refactor ll (#523) * fix bugs to former scenario * fix a bug because coding in rdloop changed * fix the bug when feedback gets no hypothesis * fix trace structure * change all trace hist when merging hypothesis to experiments * ignore some error in ruff * fix kaggle scenario bugs * refine one line * another bug * another small bug * fix ui bugs * chage kaggle train.py path --------- Co-authored-by: Xu Yang <peteryang@vip.qq.com> * fix CI * Update rdagent/app/data_science/loop.py Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com> * add samplecsv into spec prompts * fix CI --------- Co-authored-by: TPLin22 <tplin2@163.com> Co-authored-by: yuanteli <1957922024@qq.com> Co-authored-by: Xisen Wang <118058822+xisen-w@users.noreply.github.com> Co-authored-by: Bowen Xian <xianbowen@outlook.com> Co-authored-by: Xu Yang <peteryang@vip.qq.com> Co-authored-by: XianBW <36835909+XianBW@users.noreply.github.com> Co-authored-by: Tim <illking@foxmail.com> Co-authored-by: 炼金术师华华 <37462254+YeewahChan@users.noreply.github.com> Co-authored-by: Linlang <30293408+SunsetWolf@users.noreply.github.com> Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com>
164 lines
6.8 KiB
Python
164 lines
6.8 KiB
Python
from pathlib import Path
|
|
from typing import Any
|
|
|
|
import fire
|
|
|
|
from rdagent.app.data_science.conf import DS_RD_SETTING
|
|
from rdagent.components.coder.data_science.ensemble import EnsembleCoSTEER
|
|
from rdagent.components.coder.data_science.feature import FeatureCoSTEER
|
|
from rdagent.components.coder.data_science.model import ModelCoSTEER
|
|
from rdagent.components.coder.data_science.raw_data_loader import DataLoaderCoSTEER
|
|
from rdagent.components.coder.data_science.workflow import WorkflowCoSTEER
|
|
from rdagent.components.workflow.conf import BasePropSetting
|
|
from rdagent.components.workflow.rd_loop import RDLoop
|
|
from rdagent.core.exception import CoderError, RunnerError
|
|
from rdagent.core.proposal import ExperimentFeedback, HypothesisFeedback
|
|
from rdagent.core.scenario import Scenario
|
|
from rdagent.core.utils import import_class
|
|
from rdagent.log import rdagent_logger as logger
|
|
from rdagent.scenarios.data_science.dev.feedback import DSExperiment2Feedback
|
|
from rdagent.scenarios.data_science.dev.runner import DSRunner
|
|
from rdagent.scenarios.data_science.experiment.experiment import DSExperiment
|
|
from rdagent.scenarios.data_science.proposal.exp_gen import DSExpGen, DSTrace
|
|
from rdagent.scenarios.kaggle.kaggle_crawler import download_data
|
|
|
|
|
|
class DataScienceRDLoop(RDLoop):
|
|
skip_loop_error = (CoderError, RunnerError)
|
|
|
|
def __init__(self, PROP_SETTING: BasePropSetting):
|
|
logger.log_object(PROP_SETTING.competition, tag="competition")
|
|
scen: Scenario = import_class(PROP_SETTING.scen)(PROP_SETTING.competition)
|
|
|
|
### shared components in the workflow # TODO: check if
|
|
knowledge_base = (
|
|
import_class(PROP_SETTING.knowledge_base)(PROP_SETTING.knowledge_base_path, scen)
|
|
if PROP_SETTING.knowledge_base != ""
|
|
else None
|
|
)
|
|
|
|
# 1) task generation from scratch
|
|
# self.scratch_gen: tuple[HypothesisGen, Hypothesis2Experiment] = DummyHypothesisGen(scen),
|
|
|
|
# 2) task generation from a complete solution
|
|
# self.exp_gen: ExpGen = import_class(PROP_SETTING.exp_gen)(scen)
|
|
self.exp_gen = DSExpGen(scen)
|
|
self.data_loader_coder = DataLoaderCoSTEER(scen)
|
|
self.feature_coder = FeatureCoSTEER(scen)
|
|
self.model_coder = ModelCoSTEER(scen)
|
|
self.ensemble_coder = EnsembleCoSTEER(scen)
|
|
self.workflow_coder = WorkflowCoSTEER(scen)
|
|
|
|
self.runner = DSRunner(scen)
|
|
# self.summarizer: Experiment2Feedback = import_class(PROP_SETTING.summarizer)(scen)
|
|
# logger.log_object(self.summarizer, tag="summarizer")
|
|
|
|
# self.trace = KGTrace(scen=scen, knowledge_base=knowledge_base)
|
|
self.trace = DSTrace(scen=scen)
|
|
self.summarizer = DSExperiment2Feedback(scen)
|
|
super(RDLoop, self).__init__()
|
|
|
|
def direct_exp_gen(self, prev_out: dict[str, Any]):
|
|
exp = self.exp_gen.gen(self.trace)
|
|
logger.log_object(exp, tag="direct_exp_gen")
|
|
|
|
# FIXME: this is for LLM debug webapp, remove this when the debugging is done.
|
|
logger.log_object(exp, tag="debug_exp_gen")
|
|
return exp
|
|
|
|
def coding(self, prev_out: dict[str, Any]):
|
|
exp = prev_out["direct_exp_gen"]
|
|
for tasks in exp.pending_tasks_list:
|
|
exp.sub_tasks = tasks
|
|
if exp.hypothesis.component == "DataLoadSpec":
|
|
exp = self.data_loader_coder.develop(exp)
|
|
elif exp.hypothesis.component == "FeatureEng":
|
|
exp = self.feature_coder.develop(exp)
|
|
elif exp.hypothesis.component == "Model":
|
|
exp = self.model_coder.develop(exp)
|
|
elif exp.hypothesis.component == "Ensemble":
|
|
exp = self.ensemble_coder.develop(exp)
|
|
elif exp.hypothesis.component == "Workflow":
|
|
exp = self.workflow_coder.develop(exp)
|
|
else:
|
|
raise NotImplementedError(f"Unsupported component in DataScienceRDLoop: {exp.hypothesis.component}")
|
|
exp.sub_tasks = []
|
|
logger.log_object(exp, tag="coding")
|
|
return exp
|
|
|
|
def running(self, prev_out: dict[str, Any]):
|
|
exp: DSExperiment = prev_out["coding"]
|
|
if exp.next_component_required() is None:
|
|
new_exp = self.runner.run(exp)
|
|
logger.log_object(new_exp, tag="running")
|
|
return new_exp
|
|
else:
|
|
return exp
|
|
|
|
def feedback(self, prev_out: dict[str, Any]) -> ExperimentFeedback:
|
|
exp: DSExperiment = prev_out["running"]
|
|
if exp.next_component_required() is None:
|
|
feedback = self.summarizer.generate_feedback(exp, self.trace)
|
|
else:
|
|
feedback = ExperimentFeedback(
|
|
reason=f"{exp.hypothesis.component} is completed.",
|
|
decision=True,
|
|
)
|
|
logger.log_object(feedback, tag="feedback")
|
|
return feedback
|
|
|
|
def record(self, prev_out: dict[str, Any]):
|
|
e = prev_out.get(self.EXCEPTION_KEY, None)
|
|
if e is None:
|
|
self.trace.hist.append((prev_out["running"], prev_out["feedback"]))
|
|
else:
|
|
self.trace.hist.append(
|
|
(
|
|
prev_out["direct_exp_gen"] if isinstance(e, CoderError) else prev_out["coding"],
|
|
ExperimentFeedback.from_exception(e),
|
|
)
|
|
)
|
|
logger.log_object(self.trace, tag="trace")
|
|
logger.log_object(self.trace.sota_experiment(), tag="SOTA experiment")
|
|
|
|
|
|
def main(path=None, step_n=None, competition="bms-molecular-translation"):
|
|
"""
|
|
|
|
Parameters
|
|
----------
|
|
path :
|
|
path like `$LOG_PATH/__session__/1/0_propose`. It indicates that we restore the state that after finish the step 0 in loop1
|
|
step_n :
|
|
How many steps to run; if None, it will run forever until error or KeyboardInterrupt
|
|
competition :
|
|
|
|
|
|
Auto R&D Evolving loop for models in a Kaggle scenario.
|
|
You can continue running session by
|
|
.. code-block:: bash
|
|
dotenv run -- python rdagent/app/data_science/loop.py [--competition titanic] $LOG_PATH/__session__/1/0_propose --step_n 1 # `step_n` is a optional parameter
|
|
rdagent kaggle --competition playground-series-s4e8 # You are encouraged to use this one.
|
|
"""
|
|
if competition is not None:
|
|
DS_RD_SETTING.competition = competition
|
|
|
|
if DS_RD_SETTING.competition:
|
|
if DS_RD_SETTING.scen.endswith("KaggleScen"):
|
|
download_data(competition=DS_RD_SETTING.competition, settings=DS_RD_SETTING)
|
|
else:
|
|
if not Path(f"{DS_RD_SETTING.local_data_path}/{competition}").exists():
|
|
logger.error(f"Please prepare data for competition {competition} first.")
|
|
return
|
|
else:
|
|
logger.error("Please specify competition name.")
|
|
if path is None:
|
|
kaggle_loop = DataScienceRDLoop(DS_RD_SETTING)
|
|
else:
|
|
kaggle_loop = DataScienceRDLoop.load(path)
|
|
kaggle_loop.run(step_n=step_n)
|
|
|
|
|
|
if __name__ == "__main__":
|
|
fire.Fire(main)
|