Sai Charitha Akula

I am a Senior Research Scientist at Meta Superintelligence Labs, where I advance multimodal large language models through post-training techniques, with a focus on enhancing reasoning and understanding across text, image, and video modalities. Prior to Meta, I was part of Millennium's four-person core AI team, reporting to the Head of AI. There, I built enterprise-grade RAG systems and agentic workflows from the ground up and conducted research on retrieval relevance. Before that, I researched vision-centric multimodal models and diffusion models at NYU's CILVR Lab under Prof. Saining Xie. I previously spent six years as a Quantitative Researcher at Goldman Sachs' Interest Rates desk, where I built models for automated trading, pricing, and risk management.

I completed my master's in Computer Science from NYU (Courant) and bachelor's in Computer Science (Honours) from Indian Institute of Technology, Bombay.

I have served as a reviewer for NeurIPS, CVPR, and ECCV.

Outside of research, you'll usually find me exploring new cafés and nature spots, dancing, staying active, or trying out something new!

Sai Charitha Akula

Research

Broadly, my research interests lie in developing adaptive AI systems that can solve problems across diverse domains under real-world constraints. Currently, I focus on vision-language models (VLMs) and large language models (LLMs), with emphasis on efficient and scalable post-training methods across text, image, and video (RLHF, preference alignment), data-centric approaches that enhance perception and reasoning, and interpretable evaluation frameworks. I'm also excited about extending these models toward agentic behaviors through tool use and grounded interaction with structured knowledge.

Publications

Experience

Selected Projects

Website based on Jon Barron's academic website template, with personal modifications.