Publications
Also on Google Scholar.
Thesis
Granted US patents
-
2022-07Goal Oriented Dialog Generation using Dialog Template, API and Entity dataAnish Acharya, A Metallinou, T Chung, S Paul, S Chandra, C Lin, D H Tur, A MandalUS11393454 · Amazon · USPTO
-
2020-12Natural Language ProcessingCompressing word embeddings with low-rank matrix factorization for on-device natural language understanding.Anish Acharya, A Metallinou, R Goel, I DhillonUS10872601 · Amazon · USPTO
Papers
2026
-
conferenceSPLICE: Structured Prompt Local Iterative Combinatorial Evolution
Conference on Neural Information Processing Systems (NeurIPS), 2026
Treats a production prompt as a set of blocks, some frozen and some editable, and optimizes it with section-local textual gradients and elitist beam search, giving the first formal guarantees for iterative LLM-driven prompt optimization. Outperforms prior prompt optimizers across benchmarks and three model capability tiers while keeping frozen sections verbatim.
-
workshopRoPoLL: Robust Panel of LLM Judges
International Conference on Machine Learning (ICML), 2026 · AgenticUQ Workshop
Shows that averaging an LLM jury is unboundedly biased once any judge fails systematically, and replaces the average with a geometric-median aggregator that has the optimal 1/2 breakdown point. A 3-judge 38B panel beats a 675B judge by 1.31x under 30% corruption.
-
workshopOrchOpt: Validation-Gated Architecture Search for Self-Organizing Multi-Agent Orchestration
Conference on Language Modeling (COLM), 2026 · Lifelong Agents Workshop
Searches over multi-agent orchestration architectures and keeps a change only when it passes validation.
2025
-
conferenceGeometric Median Matching for Robust k-Subset Selection from Noisy Data
International Conference on Machine Learning (ICML), 2025
Selects a k-subset whose mean tracks the geometric median of noisy data, giving an O(1/k) rate under arbitrary corruption, a quadratic improvement over random sampling.
Short version: ICML 2024 FMWild Workshop
2024
-
journalNeural Distributed Source Coding
IEEE Journal on Selected Areas in Information Theory (IEEE JSAIT), 2024
Learns lossy distributed source coding with a conditional VQ-VAE, handling arbitrary correlation structure in high dimensions with state-of-the-art PSNR.
2022
-
conferenceFaster Non-Convex Federated Learning via Global and Local Momentum
Conference on Uncertainty in Artificial Intelligence (UAI), 2022 · Spotlight
Combines variance-reduced momentum at the server and at the clients to speed up non-convex federated learning under client heterogeneity and compressed communication.
-
conferenceRobust Training in High Dimensions via Block Coordinate Geometric Median Descent
International Conference on Artificial Intelligence and Statistics (AISTATS), 2022 · Spotlight
Makes geometric-median SGD practical in high dimensions by applying it to one chosen block of coordinates at a time, keeping the optimal 1/2 breakdown point for smooth non-convex problems.
Short version: Joint IFML/CCSI Symposium (Simons Institute, UC Berkeley) · Short version: NSF-TRIPODS Workshop on Communication Efficient Distributed Optimization
-
workshopPositive Unlabeled Contrastive Learning
International Conference on Machine Learning (ICML), 2022 · PODS Workshop
Extends contrastive pretraining to positive-unlabeled data with an unbiased, lower-variance loss that uses the few labeled positives.
Extended version: Understanding Contrastive Representation Learning from Positive Unlabeled (PU) Data — arXiv:2402.06038 · Short version: Simons Institute, UC Berkeley — Data-Driven Decision Processes Workshop
-
workshopLDKP: A Dataset for Identifying Keyphrases from Long Scientific Documents
ACM International Conference on Information and Knowledge Management (CIKM), 2022 · DL4SR Workshop
Releases two corpora, about 1.3M and 100K full-text scientific articles, for identifying keyphrases beyond the title and abstract.
2021
-
journalOn the Benefits of Multiple Gossip Steps in Communication Constrained Federated Learning
IEEE Transactions on Parallel and Distributed Systems (IEEE TPDS), 2021
Shows that several gossip steps between gradient updates improve decentralized learning under lossy compressed communication, even after paying for the extra communication.
-
conferenceAlexa Conversations: An Extensible Data-driven Approach for Building Task-oriented Dialogue Systems
Conference of the North American Chapter of the Association for Computational Linguistics (NAACL), 2021 · Demonstrations Track
Builds task-oriented dialogue systems from a few seed dialogues and API specifications by simulating training dialogues. The simulator improves turn-level action prediction accuracy by over 50%.
-
Conference on Empirical Methods in Natural Language Processing (EMNLP), 2021
Introduces the first summarization dataset of Hindi-English code-switched conversations: over 6,800 conversations with human-written summaries.
2019
-
conferenceOnline Embedding Compression for Text Classification using Low Rank Matrix Factorization
AAAI Conference on Artificial Intelligence (AAAI), 2019 · Spotlight
Compresses the word-embedding layer during training with low-rank factorization, reaching 90% compression with minimal accuracy loss and no added latency.
2014
-
journalExtending The Concept of Analog Butterworth Filter For Fractional Domain
Signal Processing (Elsevier), 2014
Designs fractional-order Butterworth filters in the complex w-plane, including under-, hyper-, and ultra-damped poles, and demonstrates the formulation in simulations and a frequency-domain design example.
-
Mathematics and Computers in Simulation (Elsevier), 2014
Uses real-coded genetic-algorithm optimization to tune PID controllers for four Lorenz-family multi-wing chaotic systems, with robustness checked across initial conditions.
2013
-
conferenceStability Analysis Of Delayed System Using Bode's Integral
International Conference on Computer Communication and Informatics (ICCCI), 2013
Tunes PID controllers for delayed plants using Bode's integral and a Padé approximation of the delay, then compares the approach with genetic-algorithm tuning in MATLAB simulations.
2012
-
conferenceOptimum PID Control of Multi-wing Attractors in A Family of Lorenz-like Chaotic Systems
International Conference on Computing, Communication and Networking Technologies (ICCCNT), 2012
Applies real-coded genetic-algorithm optimization to tune PID control of Lorenz-like multi-wing chaotic attractors and compares the resulting state trajectories.
-
International Conference on Computing, Communication and Networking Technologies (ICCCNT), 2012
Studies fractional-order band-pass and band-stop filters, whose characteristics are unavailable to conventional integer-order filters below second order, and optimizes their quality factor.
-
International Conference on Parallel, Distributed and Grid Computing (PDGC), 2012
Identifies nonlinear systems from outputs near several linearized operating points using feed-forward multilayer neural networks, evaluated on nuclear-reactor monitoring and AC-servo control.
2011
-
International Conference on Energy, Automation and Signal (ICEAS), 2011
Compares least-squares and instrumental-variable estimators for identifying an AC-servo position-control system under white and fractional Gaussian measurement noise.