Neel Rajani
The website template wants an address here.
But I'm not sure if I want that to be public 😅
Hi! My name is Neel, and I do research on AI Safety. I’m currently doing a PhD at the CDT for Designing Responsible NLP at the University of Edinburgh. My work focuses on the alignment of LLMs, by researching how this is typically done during post-training, why it is so brittle, and what this means for safety more broadly. AI is clearly a transformational technology, and I think that it is critical for us to get a better understanding of its guardrails and when they break.
news
| Sep 01, 2026 | Our paper on CoT Faithfulness under cue effects is up on arXiv! |
|---|---|
| Jul 27, 2026 | I’ve started a 3-month research visit at Mila in Montreal, where I’m joining Prof Siva Reddy’s lab and working on a project with the amazing Dr Verna Dankers. |
| Nov 07, 2015 | This is a placeholder announcement left over from the template that I may repurpose later... |
latest posts
| Mar 26, 2025 | a post with plotly.js |
|---|---|
| Dec 04, 2024 | a post with image galleries |
| May 14, 2024 | Google Gemini updates: Flash 1.5, Gemma 2 and Project Astra |
selected publications
- Preprint
Chain-of-Thought Faithfulness of Reasoning Models Varies with Where and How Preference Cues Are Delivered2026