RESEARCHMonitorWATCHLIST
FinEvolveBench: A Benchmark for Self-Evolving Agents on Low-Repetition Tasks with Implicit Rewards
arXiv cs.CL — Computation and Language
Factual evidence
What the source reports
Researchers introduced FinEvolveBench, a benchmark evaluating self-evolving LLM agents on non-repetitive financial data with implicit rewards.
Open source