Daisy Hill
daisyhill7
AI & ML interests
None yet
Recent Activity
upvoted a paper about 3 hours ago
S3Gym: Can LLMs Turn Self-Testing and Self-Judging into Self-Improvement? upvoted a paper about 5 hours ago
HarnessDev: Can LLMs Create and Evolve Their Own Agent Harness? upvoted a paper about 9 hours ago
StartupBench: Benchmarking General-Purpose Agents on Market-Validated End-to-End WorkflowsOrganizations
None yet