XL
lotrrrr
AI & ML interests
None yet
Recent Activity
upvoted a paper 17 days ago
WhatWorkedBench: Benchmarking Experimental Understanding in AI Agents upvoted a paper about 1 month ago
Scores Alone Do Not Prove Discovery: The Discovery Certification Protocol for Auditing AI Research Agents upvoted a paper 6 months ago
Revision or Re-Solving? Decomposing Second-Pass Gains in Multi-LLM PipelinesOrganizations
None yet