See all →

Bhagyesh Kumar profile photo

Bhagyesh Kumar

@invi-bhagyesh · AI safety

HiI am interested in prosaic alignment approaches, such as character training. Currently, I am a sophomore majoring in maths and CS, and I work at LAISR Lab, Cornell with Prof. Lionel Levine and Jonathn Chang on EigenBench,a framework for evaluating the values of LLMs.

NowI am working with Cadenza Labs on robustness of a belief in LLMs. Also, spending some time on a welfare-synthesis framework for aquaculture species with Myrias.

BeforeLast summer, I was an exchange student at USTC 🇨🇳, where I worked with Prof. Shiping Liu on Bakry-Émery curvature. Last spring, I was also a SPAR scholar, mentored by Prof. Lionel Levine and Jonathan Chang, where we worked on side effects of character training[ICML'26w]. I've also contributed to Redwood Research's Sabotage Bench.

MiscI have previously worked on adversarial attacks [AAAI'26], and Value Alignment [ICML'26]. Recently, I have been reading about formal verification for autoresearch. If you're bored, check a short explainer video I made about Higgs Boson in my high school.

You can find out what I read today at curius. Feel free to reach out and say hi at invi.bhagyesh@gmail.com

Selected research

  • Bakry–Emery Curvature Sharpness of Strongly Regular Graphs

    FuSEP'26

    Future Scientist Exchange Program, University of Science and Technology of China(Poster)

    Bhagyesh Kumar, Shiping Liu

    Poster

  • Side Effects of Character Training: Quantifying Cross-Constitution Drift in LLMs

    ICML'26

    ICML 2026 Pluralistic Alignment

    Bhagyesh Kumar, Ananya Sutradhar, Saurav Panigrahi, Jonathn Chang, Lionel Levine

    PaperSlides

  • TopoReformer: Mitigating Adversarial Attacks Using Topological Purification in OCR Models

    AAAI'26

    AAAI 2026 AI for Cyber Security(Oral)

    Bhagyesh Kumar*, A S Aravinthakshan*, Akshat Satyanarayan*, Ishaan Gakhar, Ujjwal Verma

    PaperCode

All publications →

How I fail

  • Nov 2025AAAI Undergraduate Consortium Proposal not accepted
  • Oct 2025TopoReformer(v1) rejected from NeurIPS Workshop 2025

What I'm up to now

  • Jun 2026One paper was accepted to the ICML Workshop 2026 for Pluralistic Alignment.
  • May 2026I will be spending this summer at USTC as a part of summer exchange program, working on Bakry-Émery curvature.
  • Apr 2026With my SPAR team, we released ValueArena, a leaderboard for value alignment.
  • Feb 2026Researching emergent misalignment at Cornell under SPAR
  • Jan 2026Working on evaluation awareness at Algoverse
  • Nov 2025Received $1000 scholarship to present TopoReformer at AAAI in Singapore
  • Nov 2025TopoReformer(v2) accepted at AAAI Workshop 2026