Papers.
My recent research has been focused on developing robust safety systems for large-language models (LLMs) in order to prevent catastrophic misuse of AI.
In the past, I've investigated the mechanisms that LLMs use to perform in-context learning, how to improve their capabilities, and in what ways we can better align them with our values. I've also worked on improving medical-image classification algorithms by incorporating popular computer-vision methods into the medical-image setting.
^ Blog-post only. Selected paper.
Jump to 2026 · 2025 · 2024 · 2023 · 2022 · 2021 · 2020 · 2019
Filter all · AI safety · language models · medical imaging
2026
-
* Equal contribution.
2025
-
* Equal contribution.
2024
-
* Lead contributor.
2023
2022
2021
2020
2019
Last updated September 2026.