jingyuanlm
7dd03dcac0
fix bug v6
2025-10-17 08:44:40 +00:00
jingyuanlm
8a33eb11c9
fix bug v5
2025-10-16 14:06:56 +00:00
jingyuanlm
48312032ad
fix bug v4
2025-10-15 08:34:40 +00:00
jingyuanlm
087c85e4d4
fix bug v3
2025-10-15 04:00:11 +00:00
jingyuanlm
3fe575d0c3
fix bug v2
2025-10-14 16:02:06 +00:00
jingyuanlm
63f9eb8c8e
v1
2025-10-13 10:26:18 +00:00
jingyuanlm
052ec7d321
init
2025-10-13 07:43:59 +00:00
amstrongzyf
7d749b8195
fix: fix chat_max_tokens calculation method to show true input_max_tokens ( #1241 )
...
* fix: use real max_input_length
* lint
* Update rdagent/oai/backend/litellm.py
Co-authored-by: you-n-g <you-n-g@users.noreply.github.com >
* fix lint
* lint
* lint
---------
Co-authored-by: you-n-g <you-n-g@users.noreply.github.com >
2025-09-12 15:09:20 +08:00
amstrongzyf
1265ae94e3
fix: revert 2 commits ( #1239 )
...
* revert commit:22428a45053b6eefcfb805802b8bef4384a1ddda. PR 1179
* Revert "fix: change runner prompts (#1223 )"
This reverts commit 6d3e73d679 .
* fix ruff check error
* fix: change switch place for fix_seed_and_data_split
2025-09-11 16:14:48 +08:00
you-n-g
fb628e3bca
fix: move task cancellation to finally block and fix subprocess kill typo ( #1234 )
2025-09-10 10:05:08 +08:00
you-n-g
d3d4967a12
feat: add stdout into workspace for easier debugging ( #1236 )
...
* refactor: use get_truncated_stdout for consistent stdout handling across modules
* lint
* feat: add dump_stdout_type to DSRunnerCoSTEERSettings and use in eval
* fix: avoid circular import by moving DSRunnerEvaluator import inside method
2025-09-10 10:02:49 +08:00
Tim
76b2e87348
feat: offline selector ( #1231 )
...
* offline selector test
* fix score with tensor
* sort sota list
---------
Co-authored-by: Xu Yang <peteryang@vip.qq.com >
2025-09-09 11:56:42 +08:00
Star dust
f00a5382b1
fix: add a switch for ensemble_time_upper_bound and fix some bug in main ( #1226 )
...
* change runner prompts
* v1
* ensemble_time_upper_bound
* lint
* fix inf bug in function prob_dis_torch() and ensemble prompts
* lint
* lint
---------
Co-authored-by: amstrongzyf <amstrongzyf@126.com >
2025-09-08 16:35:42 +08:00
you-n-g
c4b8baaa58
fix: increase retry count in hypothesis_gen decorator to 10 ( #1230 )
2025-09-08 12:23:58 +08:00
you-n-g
bc7a3b43b7
fix: handle ValueError in stdout shrinking and refactor shrink logic ( #1228 )
2025-09-07 23:24:45 +08:00
Tim
0a906f20e1
chore: make sure token size below limit ( #1225 )
...
* chore: make sure token size below limit
* refactor: refactor and add import
---------
Co-authored-by: amstrongzyf <amstrongzyf@126.com >
2025-09-04 15:07:42 +08:00
Star dust
6d3e73d679
fix: change runner prompts ( #1223 )
...
* change runner prompts
* v1
2025-09-04 11:31:42 +08:00
amstrongzyf
d868188f9a
fix: revert to v10 setting ( #1220 )
...
* feat: add runner_patience
* task gen prompts and torch version
* fix ensemble time bug
* small bug
* lint
* add torch
* fiix: remove useless code and change formula
* fix: rename parameter
---------
Co-authored-by: jingyuanlm <842442862@qq.com >
2025-09-03 16:37:33 +08:00
XianBW
36fec9afa6
fix: summary page bug ( #1219 )
...
* fix summary bug
* fix CI
2025-09-02 17:06:54 +08:00
you-n-g
92efe33fa9
feat: ui, support disable cache ( #1217 )
...
* fix: (drop) add cache_enable for functions in ui.utils
* fix: (drop) handle no logging, rename sota_exp to last_sota_exp for clarity in running_win
2025-09-01 15:58:05 +08:00
amstrongzyf
af9068c0b5
fix: jinja problem of enumerate ( #1216 )
2025-09-01 14:50:26 +08:00
XianBW
ea77cd6734
for ui, a simple fix ( #1215 )
2025-09-01 14:46:51 +08:00
you-n-g
1f4d190a32
fix: allow prev_out keys to be None in workspace cleanup assertion ( #1214 )
...
* fix: allow prev_out keys to be None in workspace cleanup assertion
* fix: clean workspace only for non-None DSExperiment instances
2025-08-30 07:27:12 +08:00
amstrongzyf
68af035177
fix: add missing self parameter to instance methods in DSProposalV2ExpGen ( #1213 )
2025-08-30 00:45:55 +08:00
you-n-g
68b6985191
fix: handle None output and conditional step dump in LoopBase execution ( #1212 )
2025-08-29 20:19:09 +08:00
you-n-g
05fca1acce
fix: update fallback criterion ( #1210 )
...
* fix: update fallback criterion
* fix: ensure evo_fb is initialized and used correctly in fallback logic
* refactor: rename use_new_evo to should_use_new_evo for clarity
2025-08-29 17:07:41 +08:00
you-n-g
bc3fa170b0
feat: add option to enable hyperparameter tuning only in first eval loop ( #1211 )
...
* feat: add option to enable hyperparameter tuning only in first eval loop
* fix: use total_seconds() for accurate time calculations in evolution and tracking
2025-08-29 16:59:22 +08:00
XianBW
53c2314ec0
chore: ui updates ( #1209 )
...
* show a workspace path string for easier copy, only percent existing columns when percent summary dataframe
* catch summary exception
* fix a bug
* fix ci
2025-08-29 15:05:15 +08:00
amstrongzyf
54cc2c492e
fix: fix bug for hypo_select_with_llm when not support response_schema ( #1208 )
2025-08-28 14:36:23 +08:00
XianBW
48328c5035
fix ui bug ( #1207 )
2025-08-28 14:18:58 +08:00
you-n-g
0e3a9afd0a
fix: move snapshot saving after step index update in loop execution ( #1206 )
2025-08-28 10:58:57 +08:00
amstrongzyf
1c4ab89f52
feat: enable LLM‑based hypothesis selection with time‑aware prompt & colored logging ( #1122 )
...
* add hypo select by llm (without time)
* add time_info and log color
* select no smooth
* 2 hypo
* change select
* small change
* fix bug
* fix bug and add hypothesis router and begin flag
* fix bug v1
* fix bug v2
* fix feedback
* add new model
* add filter
* fix bug v3
* fix bug v4
* change prompts v2
* fix bug v5
* fix bug v6
* fix hypo
* fix some bug(sota socre, prompts, ensemble prompts ) and add path legth.
* fix bug v7
* fix bug v8
* fix bug v9
* fix bug v10
* reset to v10 and refine
* fix: translate to english
* fix bug
* fix: use differnet increase_stage for coder & runner
* fix: use different timeout_increase_stage for coder and runner.
* fix: revert logger
* fix: remove duplicate content
* feat: implement LLM-driven extra hypothesis selection and adjust logic
* refactor: relocate Hypothesis models below TraceChallenges
* remove torch fix bug
* fix small bug
* refactor: prefix internal methods with underscore for llm-based hypothesis selection
* add fix_seed_and_data_split and enable_simple_hypothesis
* lint
* refine proposal config order
* code review; comments
* lint
* merge int to float
---------
Co-authored-by: jingyuanlm <842442862@qq.com >
Co-authored-by: Young <afe.young@gmail.com >
2025-08-27 18:36:51 +08:00
XianBW
cfb2c5fae2
chore: refine get_summary_df and sota exp stat in log utils ( #1201 )
...
* add best valid selector to get sota exp stat
* remove unused in summary & refine get_summary_df
* fix hypothesis table bug in trace
2025-08-26 18:31:21 +08:00
Linlang
8a5889fedb
style: convert CLI flags from snake_case to kebab-case ( #1198 )
...
Co-authored-by: Young <afe.young@gmail.com >
2025-08-25 20:07:01 +08:00
Xu Yang
8c62561d1c
fix: increase time default not controlled by LLM ( #1196 )
...
* fix: add longer timeout configuration for data science scenarios
* fix CI
2025-08-21 11:27:17 +08:00
XianBW
a933b6cabe
fix: kaggle competition metric direction ( #1195 )
...
* fix leaderboard and competition direction
* ui fix
2025-08-20 16:40:53 +08:00
amstrongzyf
ed5f593ae0
fix bug for embedding caused by PR 1188 ( #1194 )
2025-08-18 20:15:18 +08:00
Tim
8a0e2cf3d0
chore: refine env command ( #1193 )
2025-08-18 18:16:16 +08:00
XianBW
ad901aaf4f
fix: ui bug ( #1192 )
...
* small bug
* fix ci
2025-08-18 17:48:49 +08:00
amstrongzyf
2421fa4493
fix: enable embedding truncation ( #1188 )
...
* fix: enable embedding auto truncation
* move embedding_utils to utils.embedding
* fix bug
* fix(embedding): improve text truncation with retry strategies and binary search optimization
* new way for embedding truncate
* fix ci errors
2025-08-18 17:20:41 +08:00
xuangu-fang
bcdd957c71
feat: enable to inject diversity cross async multi-trace ( #1173 )
...
* init multi_trace_async_ diversity inject
* lint
* add task in context for diversity injection, add abstract class for diversity injection
* lint
* feat: update diversity injection strategy and enhance sibling context handling in experiment generation
* add always inject
---------
Co-authored-by: Xu Yang <peteryang@vip.qq.com >
2025-08-18 12:43:21 +08:00
you-n-g
640d64d871
change the trace number from 3 to 1 (simpler default value) ( #1191 )
2025-08-18 12:13:23 +08:00
XianBW
a18c7284a0
fix ui bugs ( #1190 )
2025-08-18 10:56:51 +08:00
you-n-g
17400a3dc4
fix: skip res_ratio check if timer or res_time is None ( #1189 )
2025-08-17 16:08:55 +08:00
Tim
eea0196f16
chore: add inference mode for model dump ( #1182 )
...
* chore: add inference mode for model dump
* check runtime environment
* update opened_trace_lines
---------
Co-authored-by: you-n-g <you-n-g@users.noreply.github.com >
2025-08-15 15:34:49 +08:00
you-n-g
5705be0512
fix: insert await asyncio.sleep(0) to yield control in loop ( #1186 )
2025-08-15 13:27:29 +08:00
you-n-g
684ef0686e
refactor: use UI_SETTING.amlt_path for combined log folder path ( #1184 )
2025-08-14 21:51:05 +08:00
Tim
d26b272cbe
chore: catch log_workflow error ( #1183 )
2025-08-14 16:12:37 +08:00
amstrongzyf
22428a4505
fix: refine prompts and add additional package info ( #1179 )
...
* refine prompts and add additional package info
* refine prompts to be specific for GBDT models
* minor refine prompts
* use include to replace duplicate info
* refine prompts
* refactor: import DSTrace from base and remove exp_gen __init__
* lint
---------
Co-authored-by: Young <afe.young@gmail.com >
2025-08-13 23:00:23 +08:00
Tim
ee2a029a00
chore: add merge statistics ( #1176 )
...
* chore: add stat for merge
* use best selector
* use trace.hist
2025-08-13 22:59:34 +08:00