AI for Science
Teaching AI to do research
Can AI do research, not just assist with it? My PhD builds LLM systems that write, run, evaluate and optimise code in a loop, then grows them into automated research workflows where experiment design, implementation and evaluation refine each other. The crux is trust: the system has to weigh scientific evidence, not just produce plausible output. My first step, ConflictBench, tests whether AI can detect, classify and explain disagreements between studies across seven research domains.





