VLDB 2026 Research / reviewers in the wild / expert
Arnab Sen Sharma
dblp:254/2046
· DBLP profile ↗
5ranked-venue papers
0as first author
5since 2021 · last 2025
—ORCID · none
Domains — the database's venue-derived domains; a paper can count in several
Artificial intelligence and machine learning · 5 · 5 since 2021
Expertise — from the expertise taxonomy: the topics of the expert's papers under the CCF categories. A weight counts papers with recency: 1 for a paper about the topic, 0.3 when the topic is its context, halved every five years.
| Artificial intelligence
4 papers |
Trustworthy machine learning · 28% Language models and text generation · 21% Efficient and distributed learning · 14% |
Topics — the 9 heaviest of 12, each with the papers that count most for it
| Topic | Weight | Papers | Last | Evidence papers |
|---|---|---|---|---|
Machine learning › Efficient and distributed learning
distributed training |
0.9 | 1 | 2025 | NNsight and NDIF: Democratizing Access to Open-Weight Foundation Model Internals · ICLR 2025 |
Machine learning › Trustworthy machine learning › interpretability › mechanistic interpretability
function vectors |
0.8 | 1 | 2024 | Function Vectors in Large Language Models · ICLR 2024 |
Natural language and speech › Language models and text generation
in-context learning |
0.8 | 1 | 2024 | Function Vectors in Large Language Models · ICLR 2024 |
Machine learning › Trustworthy machine learning
interpretability |
0.8 | 1 | 2024 | Linearity of Relation Decoding in Transformer Language Models · ICLR 2024 |
Machine learning › Trustworthy machine learning › interpretability
mechanistic interpretability |
0.8 | 1 | 2024 | Function Vectors in Large Language Models · ICLR 2024 |
Machine learning › Probabilistic and Bayesian machine learning › causal inference › causal effect estimation
mediation analysis |
0.8 | 1 | 2024 | Function Vectors in Large Language Models · ICLR 2024 |
Natural language and speech › Language models and text generation
knowledge editing |
0.7 | 1 | 2023 | Mass-Editing Memory in a Transformer · ICLR 2023 |
Machine learning › Efficient and distributed learning
inference serving |
0.3 | 1 | 2025 | NNsight and NDIF: Democratizing Access to Open-Weight Foundation Model Internals · ICLR 2025 |
Natural language and speech › Language models and text generation
large language model |
0.3 | 1 | 2025 | NNsight and NDIF: Democratizing Access to Open-Weight Foundation Model Internals · ICLR 2025 |
Methods — techniques the papers use, named apart from their topics
intervention graph · 0.9deferred remote execution · 0.9probing · 0.8linear transformation · 0.8causal mediation analysis · 0.8rank-one update · 0.7model editing · 0.7
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2025 | NNsight and NDIF: Democratizing Access to Open-Weight Foundation Model InternalsabstractWe introduce NNsight and NDIF, technologies that work in tandem to enable scientific study of the representations and computations learned by very large neural networks. NNsight is an open-source system that extends PyTorch to introduce deferred remote execution. The National Deep Inference Fabric (NDIF) is a scalable inference service that executes NNsight requests, allowing users to share GPU resources and pretrained models. These technologies are enabled by the Intervention Graph, an architecture developed to decouple experimental design from model runtime. Together, this framework provides transparent and efficient access to the internals of deep neural networks such as very large language models (LLMs) without imposing the cost or complexity of hosting customized models individually. We conduct a quantitative survey of the machine learning literature that reveals a growing gap in the study of the internals of large-scale AI. We demonstrate the design and use of our framework to address this gap by enabling a range of research methods on huge models. Finally, we conduct benchmarks to compare performance with previous approaches.
Code, documentation, and tutorials are available at https://nnsight.net/. Jaden Fiotto-Kaufman, Alexander R. Loftus, Eric Todd, Jannik Brinkmann, Koyena Pal, Dmitrii Troitskii, Michael Ripa, Adam Belfki, Can Rager, Caden Juang, Aaron Mueller, Samuel Marks, Arnab Sen Sharma, Francesca Lucchetti, Nikhil Prakash, Carla E. Brodley, Arjun Guha, Jonathan Bell 0001, Byron C. Wallace, David Bau |
ICLR | 13 |
| 2024 | Linearity of Relation Decoding in Transformer Language ModelsabstractMuch of the knowledge encoded in transformer language models (LMs) may be expressed in terms of relations: relations between words and their synonyms, entities and their attributes, etc. We show that, for a subset of relations, this computation is well-approximated by a single linear transformation on the subject representation. Linear relation representations may be obtained by constructing a first-order approximation to the LM from a single prompt, and they exist for a variety of factual, commonsense, and linguistic relations. However, we also identify many cases in which LM predictions capture relational knowledge accurately, but this knowledge is not linearly encoded in their representations. Our results thus reveal a simple, interpretable, but heterogeneously deployed knowledge representation strategy in transformer LMs. Evan Hernandez, Arnab Sen Sharma, Tal Haklay, Kevin Meng, Martin Wattenberg, Jacob Andreas, Yonatan Belinkov, David Bau |
ICLR | 2 |
| 2024 | Function Vectors in Large Language ModelsabstractWe report the presence of a simple neural mechanism that represents an input-output function as a vector within autoregressive transformer language models (LMs). Using causal mediation analysis on a diverse range of in-context-learning (ICL) tasks, we find that a small number attention heads transport a compact representation of the demonstrated task, which we call a function vector (FV). FVs are robust to changes in context, i.e., they trigger execution of the task on inputs such as zero-shot and natural text settings that do not resemble the ICL contexts from which they are collected. We test FVs across a range of tasks, models, and layers and find strong causal effects across settings in middle layers. We investigate the internal structure of FVs and find while that they often contain information that encodes the output space of the function, this information alone is not sufficient to reconstruct an FV. Finally, we test semantic vector composition in FVs, and find that to some extent they can be summed to create vectors that trigger new complex tasks. Our findings show that compact, causal internal vector representations of function abstractions can be explicitly extracted from LLMs. Eric Todd, Millicent Li, Arnab Sen Sharma, Aaron Mueller, Byron C. Wallace, David Bau |
ICLR | 3 |
| 2023 | Mass-Editing Memory in a Transformer
Kevin Meng, Arnab Sen Sharma, Alex Andonian, Yonatan Belinkov, David Bau |
ICLR | 2 |
| 2022 | BD-SHS: A Benchmark Dataset for Learning to Detect Online Bangla Hate Speech in Different Social ContextsabstractSocial media platforms and online streaming services have spawned a new breed of Hate Speech (HS). Due to the massive amount of user-generated content on these sites, modern machine learning techniques are found to be feasible and cost-effective to tackle this problem. However, linguistically diverse datasets covering different social contexts in which offensive language is typically used are required to train generalizable models. In this paper, we identify the shortcomings of existing Bangla HS datasets and introduce a large manually labeled dataset BD-SHS that includes HS in different social contexts. The labeling criteria were prepared following a hierarchical annotation process, which is the first of its kind in Bangla HS to the best of our knowledge. The dataset includes more than 50,200 offensive comments crawled from online social networking sites and is at least 60% larger than any existing Bangla HS datasets. We present the benchmark result of our dataset by training different NLP models resulting in the best one achieving an F1-score of 91.0%. In our experiments, we found that a word embedding trained exclusively using 1.47 million comments from social media and streaming sites consistently resulted in better modeling of HS detection in comparison to other pre-trained embeddings. Our dataset and all accompanying codes is publicly available at github.com/naurosromim/hate-speech-dataset-for-Bengali-social-media Nauros Romim, Mosahed Ahmed, Arnab Sen Sharma, Hriteshwar Talukder, Mohammad Ruhul Amin |
LREC | 4 |