All papers
Every paper that carries a page of its own here. The Papers page leads with what is trending and lets you filter the recent catalog; this is the plain index of the rest.
The Portiloop: a deep learning-based open science tool for closed-loop brain stimulationarXiv 2021Autoencoding a Soft Touch to Learn Grasping from On-land to UnderwaterarXiv 2023Does Self-supervised Learning Really Improve Reinforcement Learning from Pixels?arXiv 2022Distributionally Robust Recourse Actiondistributionally-robust-recourse-actionBioFusionNet: Deep Learning-Based Survival Risk Stratification in ER+ Breast Cancer Through Multifeature and Multimodal Data FusionarXiv 2024A Law of Next-Token Prediction in Large Language ModelsarXiv 2024Interpretable structural model error discovery from sparse assimilation
increments using spectral bias-reduced neural networks: A quasi-geostrophic
turbulence test casearXiv 2023Curriculum Direct Preference Optimization for Diffusion and Consistency ModelsCVPR 2025 1AutoHall: Automated Hallucination Dataset Generation for Large Language ModelsarXiv 2023Simplifying Momentum-based Positive-definite Submanifold Optimization with Applications to Deep LearningarXiv 2023Rapid Adaptation in Online Continual Learning: Are We Evaluating It Right?ICCV 2023 1PEACE: Cross-Platform Hate Speech Detection- A Causality-guided FrameworkarXiv 2023Probing Language Models on Their Knowledge SourcearXiv 2024Polymath: A Challenging Multi-modal Mathematical Reasoning BenchmarkarXiv 2024E2S2: Encoding-Enhanced Sequence-to-Sequence Pretraining for Language Understanding and GenerationarXiv 2022KidLM: Advancing Language Models for Children -- Early Insights and Future DirectionsarXiv 2024DOTS: Learning to Reason Dynamically in LLMs via Optimal Reasoning Trajectories SearcharXiv 2024Practical Region-level Attack against Segment Anything ModelsarXiv 2024A Cascade Approach to Neural Abstractive Summarization with Content Selection and FusionAsian Chapter of the Association for Computational Linguistics 2020HELPD: Mitigating Hallucination of LVLMs by Hierarchical Feedback Learning with Vision-enhanced Penalty DecodingarXiv 2024Improving the Shortest Plank: Vulnerability-Aware Adversarial Training for Robust Recommender SystemarXiv 2024Improved Pothole Detection Using YOLOv7 and ESRGANarXiv 2023CulFiT: A Fine-grained Cultural-aware LLM Training Paradigm via Multilingual Critique Data SynthesisarXiv 2025LLMs in Biomedicine: A study on clinical Named Entity RecognitionarXiv 2024MAQA: Evaluating Uncertainty Quantification in LLMs Regarding Data UncertaintyarXiv 2024Objects Can Move: 3D Change Detection by Geometric Transformation ConstistencyarXiv 2022Specialized Foundation Models Struggle to Beat Supervised BaselinesarXiv 2024NLI Data Sanity Check: Assessing the Effect of Data Corruption on Model PerformanceNoDaLiDa 2021 5Evaluating Deep Graph Neural Networksevaluating-deep-graph-neural-networks-1Effective Clustering on Large Attributed Bipartite GraphsarXiv 2024Synthetic Dialogue Dataset Generation using LLM AgentsarXiv 2024Black-box Unsupervised Domain Adaptation with Bi-directional Atkinson-Shiffrin MemoryICCV 2023 1Nonparametric Teaching of Implicit Neural RepresentationsarXiv 2024Emergence of a High-Dimensional Abstraction Phase in Language TransformersarXiv 2024Improving Generative Model-based Unfolding with Schrödinger BridgesarXiv 2023Data Contamination Through the Lens of TimearXiv 2023VideoUFO: A Million-Scale User-Focused Dataset for Text-to-Video GenerationarXiv 2025Annotator-Centric Active Learning for Subjective NLP TasksarXiv 2024What do tokens know about their characters and how do they know it?NAACL 2022 7UNKs Everywhere: Adapting Multilingual Language Models to New ScriptsEMNLP 2021 11EchoPrompt: Instructing the Model to Rephrase Queries for Improved In-context LearningarXiv 2023Evaluating Open-Domain Dialogues in Latent Space with Next Sentence Prediction and Mutual InformationarXiv 2023Discovery of interpretable structural model errors by combining Bayesian sparse regression and data assimilation: A chaotic Kuramoto-Sivashinsky test casearXiv 2021This is not correct! Negation-aware Evaluation of Language Generation SystemsarXiv 2023MaiBaam: A Multi-Dialectal Bavarian Universal Dependency TreebankarXiv 2024An Empirical Study on Developers Shared Conversations with ChatGPT in
GitHub Pull Requests and IssuesarXiv 2024Beyond Good Intentions: Reporting the Research Landscape of NLP for Social GoodarXiv 2023Cheetah: Natural Language Generation for 517 African LanguagesarXiv 2024Multilinear Operator NetworksarXiv 2024Are Large Language Models Good at Utility Judgments?arXiv 2024Hierarchical Catalogue Generation for Literature Review: A BenchmarkarXiv 2023Model-Agnostic Gender Debiased Image CaptioningCVPR 2023 1Unveiling LLMs: The Evolution of Latent Representations in a Dynamic
Knowledge GrapharXiv 2024AI-Assisted Generation of Difficult Math QuestionsarXiv 2024SweEval: Do LLMs Really Swear? A Safety Benchmark for Testing Limits for Enterprise UsearXiv 2025On the Relationship between Truth and Political Bias in Language ModelsarXiv 2024SAEs Are Good for Steering -- If You Select the Right FeaturesarXiv 2025Continual Learning with Dynamic Sparse Training: Exploring Algorithms for Effective Model UpdatesarXiv 2023Advancing Regular Language Reasoning in Linear Recurrent Neural NetworksarXiv 2023Traces of Memorisation in Large Language Models for CodearXiv 2023Backdoor Secrets Unveiled: Identifying Backdoor Data with Optimized Scaled Prediction ConsistencyarXiv 2024RepLiQA: A Question-Answering Dataset for Benchmarking LLMs on Unseen Reference ContentarXiv 2024COVR: A test-bed for Visually Grounded Compositional Generalization with real imagesEMNLP 2021 11Who Wrote This? The Key to Zero-Shot LLM-Generated Text Detection Is GECScorearXiv 2024Beyond Attentive Tokens: Incorporating Token Importance and Diversity for Efficient Vision TransformersCVPR 2023 1A Theory of Unsupervised Translation Motivated by Understanding Animal Communicationa-theory-of-unsupervised-translationVisually-Aware Context Modeling for News Image CaptioningarXiv 2023UNK-VQA: A Dataset and a Probe into the Abstention Ability of Multi-modal Large ModelsarXiv 2023A Large-Scale Empirical Study on Improving the Fairness of Image Classification ModelsarXiv 2024Curriculum Dataset DistillationarXiv 2024BabyStories: Can Reinforcement Learning Teach Baby Language Models to Write Better Stories?arXiv 2023F4Splat: Feed-Forward Predictive Densification for Feed-Forward 3D Gaussian SplattingarXiv 2026CoDe: Blockwise Control for Denoising Diffusion ModelsarXiv 2025Event Transition Planning for Open-ended Text GenerationFindings (ACL) 2022 5Improving Personality Consistency in Conversation by Persona ExtendingarXiv 2022Learning Robot Manipulation from Cross-Morphology DemonstrationarXiv 2023WinoGAViL: Gamified Association Benchmark to Challenge Vision-and-Language ModelsarXiv 2022RAFT: Realistic Attacks to Fool Text DetectorsarXiv 2024To Trust or Not To Trust Prediction Scores for Membership Inference AttacksarXiv 2021Automating Human Tutor-Style Programming Feedback: Leveraging GPT-4 Tutor Model for Hint Generation and GPT-3.5 Student Model for Hint ValidationarXiv 2023GOAt: Explaining Graph Neural Networks via Graph Output AttributionarXiv 2024AutoBench-V: Can Large Vision-Language Models Benchmark Themselves?arXiv 2024Algorithm Selection for Deep Active Learning with Imbalanced Datasetsalgorithm-selection-for-deep-active-learningAn Empirical Study on Cross-lingual Vocabulary Adaptation for Efficient Language Model InferencearXiv 2024Fast Rates for Maximum Entropy ExplorationarXiv 2023Playing Along: Learning a Double-Agent Defender for Belief Steering via Theory of MindarXiv 2026Explaining Text Similarity in Transformer ModelsarXiv 2024Experimenting with Emerging RISC-V Systems for Decentralised Machine LearningarXiv 2023Synthesis of Batik Motifs using a Diffusion -- Generative Adversarial NetworkarXiv 2023Stable Mean Teacher for Semi-supervised Video Action DetectionarXiv 2024On the Language Neutrality of Pre-trained Multilingual RepresentationsFindings of the Association for Computational Linguistics 2020Explaining Speech Classification Models via Word-Level Audio Segments and Paralinguistic FeaturesarXiv 2023Generalized Funnelling: Ensemble Learning and Heterogeneous Document Embeddings for Cross-Lingual Text ClassificationarXiv 2021Refiner: Restructure Retrieval Content Efficiently to Advance Question-Answering CapabilitiesarXiv 2024Exact Gauss-Newton Optimization for Training Deep Neural NetworksarXiv 2024Localizing Active Objects from Egocentric Vision with Symbolic World KnowledgearXiv 2023Top General Performance = Top Domain Performance? DomainCodeBench: A Multi-domain Code Generation BenchmarkarXiv 2024World2Minecraft: Occupancy-Driven Simulated Scenes ConstructionarXiv 2026On the Limit of Language Models as Planning FormalizersarXiv 2024PLeaS -- Merging Models with Permutations and Least SquaresarXiv 2024Humor in AI: Massive Scale Crowd-Sourced Preferences and Benchmarks for Cartoon CaptioningarXiv 2024QACE: Asking Questions to Evaluate an Image CaptionFindings (EMNLP) 2021 11Monolingual and Cross-Lingual Acceptability Judgments with the Italian CoLA corpusFindings (EMNLP) 2021 11GenCodeSearchNet: A Benchmark Test Suite for Evaluating Generalization in Programming Language UnderstandingarXiv 2023V_{0.5}: Generalist Value Model as a Prior for Sparse RL RolloutsarXiv 2026Clover: Regressive Lightweight Speculative Decoding with Sequential KnowledgearXiv 2024Towards Nonlinear-Motion-Aware and Occlusion-Robust Rolling Shutter CorrectionICCV 2023 1Machine Translation Hallucination Detection for Low and High Resource Languages using Large Language ModelsarXiv 2024Accelerating Data Generation for Neural Operators via Krylov Subspace RecyclingarXiv 2024Aligning Language Models to Explicitly Handle AmbiguityarXiv 2024Query Understanding via Intent Description GenerationarXiv 2020How Predictable Are Large Language Model Capabilities? A Case Study on BIG-bencharXiv 2023Layer-Level Self-Exposure and Patch: Affirmative Token Mitigation for Jailbreak Attack DefensearXiv 2025An Exploratory Study on Fine-Tuning Large Language Models for Secure
Code GenerationarXiv 2024MetaphorVU: Towards Metaphorical Video UnderstandingarXiv 2026Loss-to-Loss Prediction: Scaling Laws for All DatasetsarXiv 2024Chinese Fine-Grained Financial Sentiment Analysis with Large Language ModelsarXiv 2023ClusT3: Information Invariant Test-Time Trainingclust3-information-invariant-test-timeDefining and Extracting generalizable interaction primitives from DNNsarXiv 2024A Comprehensive Analysis of the Effectiveness of Large Language Models as Automatic Dialogue EvaluatorsarXiv 2023Revisiting Instruction Fine-tuned Model Evaluation to Guide Industrial ApplicationsarXiv 2023Label Dependent Attention Model for Disease Risk Prediction Using Multimodal Electronic Health RecordsarXiv 2022Bridging the Gap between Synthetic and Authentic Images for Multimodal Machine TranslationarXiv 2023Statistical mechanics of continual learning: variational principle and mean-field potentialarXiv 2022World to Code: Multi-modal Data Generation via Self-Instructed Compositional Captioning and FilteringarXiv 2024A Learnable Prior Improves Inverse Tumor Growth ModelingarXiv 2024MMCert: Provable Defense against Adversarial Attacks to Multi-modal ModelsCVPR 2024 1GRAFENNE: Learning on Graphs with Heterogeneous and Dynamic Feature SetsarXiv 2023ProBench: Benchmarking Large Language Models in Competitive ProgrammingarXiv 2025The language of prompting: What linguistic properties make a prompt successful?arXiv 2023Diffusion Models as Artists: Are we Closing the Gap between Humans and Machines?arXiv 2023From Loops to Oops: Fallback Behaviors of Language Models Under UncertaintyarXiv 2024DoG-Instruct: Towards Premium Instruction-Tuning Data via Text-Grounded Instruction WrappingarXiv 2023Understanding the Role of Mixup in Knowledge Distillation: An Empirical StudyarXiv 2022Perplexity Trap: PLM-Based Retrievers Overrate Low Perplexity DocumentsarXiv 2025Multilingual Text-to-Image Generation Magnifies Gender Stereotypes and Prompt Engineering May Not Help YouarXiv 2024CultureBERT: Measuring Corporate Culture With Transformer-Based Language ModelsarXiv 2022ChemTEB: Chemical Text Embedding Benchmark, an Overview of Embedding Models Performance & Efficiency on a Specific DomainarXiv 2024Improving Multimodal Learning with Multi-Loss Gradient ModulationarXiv 2024Fairness-Aware Structured Pruning in TransformersarXiv 2023Do Parameters Reveal More than Loss for Membership Inference?arXiv 2024Neural Markov Jump ProcessesarXiv 2023IGA : An Intent-Guided Authoring AssistantarXiv 2021On Sequential Bayesian Inference for Continual LearningarXiv 2023Inference Scaling fLaws: The Limits of LLM Resampling with Imperfect VerifiersarXiv 2024Critical Learning Periods Emerge Even in Deep Linear NetworksarXiv 2023LLM-enhanced Self-training for Cross-domain Constituency ParsingarXiv 2023Taming Knowledge Conflicts in Language ModelsarXiv 2025PMIndiaSum: Multilingual and Cross-lingual Headline Summarization for Languages in IndiaarXiv 2023Look at the Text: Instruction-Tuned Language Models are More Robust Multiple Choice Selectors than You ThinkarXiv 2024Back Transcription as a Method for Evaluating Robustness of Natural Language Understanding Models to Speech Recognition ErrorsarXiv 2023Evaluating Large Language Models in Semantic Parsing for Conversational Question Answering over Knowledge GraphsarXiv 2024Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language ModelsarXiv 2024PentestGPT: An LLM-empowered Automatic Penetration Testing ToolarXiv 2023On the Loss of Context-awareness in General Instruction Fine-tuningarXiv 2024Lexical Generalization Improves with Larger Models and Longer TrainingarXiv 2022Mirage: Model-Agnostic Graph Distillation for Graph ClassificationarXiv 2023Counterfactual Identifiability of Bijective Causal ModelsarXiv 2023Investigating the Impact of Direct Punishment on the Emergence of Cooperation in Multi-Agent Reinforcement Learning SystemsarXiv 2023Adaptive Federated Learning with Auto-Tuned ClientsarXiv 2023PRISM: Position-encoded Regressive Inverse Spectral Model for Multilayer Thin-Film DesignarXiv 2026Mind the Memory Gap: Unveiling GPU Bottlenecks in Large-Batch LLM InferencearXiv 2025CRONOS: Benchmarking Counterfactual Physical Consistency in Video ModelsarXiv 2026Functional Bayesian Tucker Decomposition for Continuous-indexed Tensor DataarXiv 2023Game Plot Design with an LLM-powered Assistant: An Empirical Study with Game DesignersarXiv 2024Position Aware 60 GHz mmWave Beamforming for V2V Communications Utilizing Deep LearningarXiv 2024ID.8: Co-Creating Visual Stories with Generative AIarXiv 2023Did You Really Just Have a Heart Attack? Towards Robust Detection of Personal Health Mentions in Social MediaarXiv 2018Instances Need More Care: Rewriting Prompts for Instances with LLMs in the Loop Yields Better Zero-Shot PerformancearXiv 2023Comparing Self-Supervised Learning Models Pre-Trained on Human Speech and Animal Vocalizations for Bioacoustics ProcessingarXiv 2025PoliTune: Analyzing the Impact of Data Selection and Fine-Tuning on Economic and Political Biases in Large Language ModelsarXiv 2024What do you Mean? The Role of the Mean Function in Bayesian OptimisationarXiv 2020Investigating Zero-Shot Generalizability on Mandarin-English Code-Switched ASR and Speech-to-text Translation of Recent Foundation Models with Self-Supervision and Weak SupervisionarXiv 2023Profiling Neural Blocks and Design Spaces for Mobile Neural Architecture SearcharXiv 2021Breast Tumor Classification Using EfficientNet Deep Learning ModelarXiv 2024Paraphrase Detection: Human vs. Machine ContentarXiv 2023CoMT: Chain-of-Medical-Thought Reduces Hallucination in Medical Report GenerationarXiv 2024Florence: A New Foundation Model for Computer VisionarXiv 2021Multi-Peptide: Multimodality Leveraged Language-Graph Learning of Peptide PropertiesarXiv 2024Pre-Trained Language-Meaning Models for Multilingual Parsing and GenerationarXiv 2023Differentially Private SGD Without Clipping Bias: An Error-Feedback ApproacharXiv 2023IMDB-WIKI-SbS: An Evaluation Dataset for Crowdsourced Pairwise
ComparisonsarXiv 2021How sensitive are translation systems to extra contexts? Mitigating gender bias in Neural Machine Translation models through relevant contextsarXiv 2022Mismatch Quest: Visual and Textual Feedback for Image-Text MisalignmentarXiv 2023Vector-Valued Control VariatesarXiv 2021Text Generation: A Systematic Literature Review of Tasks, Evaluation, and ChallengesarXiv 2024Sentiment-enhanced Graph-based Sarcasm Explanation in DialoguearXiv 2024Using Language Model to Bootstrap Human Activity Recognition Ambient Sensors Based in Smart HomesarXiv 2021Your Finetuned Large Language Model is Already a Powerful Out-of-distribution DetectorarXiv 2024Jointly-Learned Exit and Inference for a Dynamic Neural Network : JEI-DNNarXiv 2023Polyglot-Lion: Efficient Multilingual ASR for Singapore via Balanced Fine-Tuning of Qwen3-ASRarXiv 2026Synchronous Bidirectional Learning for Multilingual Lip ReadingarXiv 2020Evaluating Gender Bias in Natural Language Inferenceevaluating-gender-bias-in-natural-languageNot All Metrics Are Guilty: Improving NLG Evaluation by Diversifying ReferencesarXiv 2023Liar, Liar, Logical Mire: A Benchmark for Suppositional Reasoning in Large Language ModelsarXiv 2024SA-VLA: Spatially-Aware Flow-Matching for Vision-Language-Action Reinforcement LearningarXiv 2026Lived Experience Not Found: LLMs Struggle to Align with Experts on Addressing Adverse Drug Reactions from Psychiatric Medication UsearXiv 2024DreamStruct: Understanding Slides and User Interfaces via Synthetic Data GenerationarXiv 2024AgAsk: An Agent to Help Answer Farmer's Questions From Scientific DocumentsarXiv 2022Language Model Behavior: A Comprehensive SurveyarXiv 2023