AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
18 May 2026

DetectRL-X: Towards Reliable Multilingual and Real-World LLM-Generated Text Detection

Model ReleasesDGX agent

arXiv:2605.15518v1 Announce Type: new Abstract: The effective detection and governance of Large Language Model (LLM) generated content has become increasingly critical due to the growing risk of misus

DimMem: Dimensional Structuring for Efficient Long-Term Agent Memory

Model ReleasesDGX agent

arXiv:2605.15759v1 Announce Type: new Abstract: Large language model (LLM) agents require long-term memory to leverage information from past interactions. However, existing memory systems often face a

DiscoExplorer: An Open Interface for the Study of Multilingual Discourse Relations

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.15304v1 Announce Type: new Abstract: The relations connecting propositions in discourse such as cause (A because B) or concession (A although B) are a subject of intense interest in Computa

DiscussLLM: Teaching Large Language Models When to Speak

TutorialsDGX agent

arXiv:2508.18167v2 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities in understanding and generating human-like text, yet they largely operate as

Dynamic Chunking for Diffusion Language Models

Model ReleasesDGX agent

arXiv:2605.15676v1 Announce Type: new Abstract: Block discrete diffusion language models factorize a sequence autoregressively over fixed-size positional blocks, decoupling within-block parallel denoi

Eskwai for Students: Generative AI Assistant for Legal Education in Ghana

SafetyDGX agent

arXiv:2605.15380v1 Announce Type: new Abstract: Recent advances in generative AI have shown their potential to be leveraged for legal education. Yet, work on the development and deployment of such sys

Evaluating Chinese Ambiguity Understanding in Large Language Models

Model ReleasesDGX agent

arXiv:2605.15635v1 Announce Type: new Abstract: Linguistic ambiguity is critical to the robustness of Large Language Models (LLMs), yet existing research focuses mostly on English, with limited attent

Few-Shot Large Language Models for Actionable Triage Categorization of Online Patient Inquiries

Model ReleasesDGX agent

arXiv:2605.15680v1 Announce Type: new Abstract: Online patient inquiries are often informal, incomplete, and written before professional assessment, yet they must still be routed to an appropriate lev

Few-Step Diffusion Language Models via Trajectory Self-Distillation

ResearchDGX agent

arXiv:2602.12262v3 Announce Type: replace Abstract: Diffusion large language models (DLLMs) have emerged as powerful generative models with the promise of fast text generation through parallel decodin

FINESSE-Bench: A Hierarchical Benchmark Suite for Financial Domain Knowledge and Technical Analysis in Large Language Models

Model ReleasesDGX agent

arXiv:2605.15482v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly being applied to financial analysis, reporting, investment decision support, risk management, compliance,

FinReporting: An Agentic Workflow for Localized Reporting of Cross-Jurisdiction Financial Disclosures

SafetyDGX agent

arXiv:2604.05966v2 Announce Type: replace Abstract: Financial reporting systems increasingly leverage Large Language Models (LLMs) to extract and summarize corporate disclosures. However, most existin

Fluency and Faithfulness in Human and Machine Literary Translation

ResearchDGX agent

arXiv:2605.15282v1 Announce Type: new Abstract: Literary translation requires balancing target-language fluency with faithfulness to the source. Recent large language models (LLMs) often produce fluen

ForMaT: Dataset for Visually-Grounded Multilingual PDF Translation

Model ReleasesDGX agent

arXiv:2605.15794v1 Announce Type: new Abstract: We present ForMaT (Format-Preserving Multilingual Translation), a parallel corpus of 3,956 PDFs across 15 language pairs that preserves original layout

GiLT: Augmenting Transformer Language Models with Dependency Graphs

Model ReleasesDGX agent

arXiv:2605.15562v1 Announce Type: new Abstract: Augmenting Transformers with linguistic structures effectively enhances the syntactic generalization performance of language models. Previous work in th

Greedy or not, here I come: Language production under vocabulary constraints in humans and resource-rational models

ApplicationsDGX agent

arXiv:2605.15365v1 Announce Type: new Abstract: Communicating using only a limited vocabulary is a common but challenging cognitive phenomenon, requiring an ideal communicator to plan carefully to opt

Hallucinations are inevitable but can be made statistically negligible

ResearchDGX agent

arXiv:2502.12187v3 Announce Type: replace Abstract: Hallucinations, a phenomenon where a language model (LM) generates nonfactual content, pose a significant challenge to the practical deployment of L

Improving Cross-Cultural Survey Simulation with Calibrated Value Personas

ResearchDGX agent

arXiv:2605.16193v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used to simulate human opinions and survey responses, but their ability to reproduce population responses

Introducing MELI: the Mandarin-English Language Interview Corpus

Model ReleasesDGX agent

arXiv:2603.27043v2 Announce Type: replace Abstract: We introduce the Mandarin-English Language Interview (MELI) Corpus, an open-source resource of 29.8 hours of speech from 51 Mandarin-English bilingu

Judge Circuits

Model ReleasesDGX agent

arXiv:2605.16023v1 Announce Type: new Abstract: LLM-as-a-judge has become the dominant paradigm for grading model outputs at scale, yet the same model assigns systematically different scores when its

Linked Multi-Model Data on Russian Domestic and Foreign Policy Speeches

SafetyDGX agent

arXiv:2605.15886v1 Announce Type: new Abstract: This paper introduces a dataset of interlinked multimodal political communications from the Russian government, addressing persistent deficiencies in th

Measuring Maximum Activations in Open Large Language Models

Model ReleasesDGX agent

arXiv:2605.15572v1 Announce Type: new Abstract: The dynamic range of activations is a first-order constraint for low-bit quantization, activation scaling, and stable LLM inference. Prior work characte

MHGraphBench: Knowledge Graph-Grounded Benchmarking of Mental Health Knowledge in Large Language Models

Model ReleasesDGX agent

arXiv:2605.15589v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used in the mental health domain, yet it remains unclear how well they capture related biomedical knowledg

Mind the Motions: Benchmarking Theory-of-Mind in Everyday Body Language

ResearchDGX agent

arXiv:2511.15887v2 Announce Type: replace Abstract: Our ability to interpret others' mental states through nonverbal cues (NVCs) is fundamental to our survival and social cohesion. While existing Theo

Multi-Level Contextual Token Relation Modeling for Machine-Generated Text Detection

Local AiDGX agent

arXiv:2605.16107v1 Announce Type: new Abstract: Machine-generated texts (MGTs) pose risks such as disinformation and phishing, underscoring the need for reliable detection. Metric-based methods, which

Neural Activation Patterns Across Language Model Architectures: A Comprehensive Analysis of Cognitive Task Performance

ResearchDGX agent

arXiv:2605.15436v1 Announce Type: new Abstract: This paper presents a comprehensive analysis of neural activation patterns across six distinct large language model (LLM) architectures, examining their

Optimized Three-Dimensional Photovoltaic Structures with LLM guided Tree Search

AgentsDGX agent

arXiv:2605.16191v1 Announce Type: new Abstract: We present a case study for how AI coding systems can be used to generate novel scientific hypotheses. We combine a generic coding agent (Google's AntiG

PerfCodeBench: Benchmarking LLMs for System-Level High-Performance Code Optimization

Model ReleasesDGX agent

arXiv:2605.15222v1 Announce Type: cross Abstract: Large language models (LLMs) can often generate functionally correct code, but their ability to produce efficient implementations for performance-crit

Prompt Stability Scoring for Text Annotation with Large Language Models

ResearchDGX agent

arXiv:2407.02039v3 Announce Type: replace Abstract: Researchers are increasingly using language models (LMs) for text annotation. These approaches rely only on a prompt telling the model to return a g

PSD: Pushing the Pareto Frontier of Diffusion LLMs via Parallel Speculative Decoding

SafetyDGX agent

arXiv:2605.15609v1 Announce Type: new Abstract: Diffusion large language models (dLLMs) generate text by iteratively denoising masked token sequences. Although dLLMs can predict all masked positions i

RapidUn: Influence-Driven Parameter Reweighting for Efficient Large Language Model Unlearning

Model ReleasesDGX agent

arXiv:2512.04457v2 Announce Type: replace Abstract: Removing specific data influence from large language models (LLMs) remains challenging, as retraining is costly and existing approximate unlearning

Reasoning Models Don't Just Think Longer, They Move Differently

ResearchDGX agent

arXiv:2605.15454v1 Announce Type: new Abstract: Reasoning-trained language models often spend more tokens on harder problems, but longer chains of thought do not show whether a model is merely computi

Reference Games as a Testbed for the Alignment of Model Uncertainty and Clarification Requests

SafetyDGX agent

arXiv:2601.07820v2 Announce Type: replace Abstract: In human conversation, both interlocutors play an active role in maintaining mutual understanding. When listeners are uncertain about what speakers

Response-Conditioned Parallel-to-Sequential Orchestration for Multi-Agent Systems

SafetyDGX agent

arXiv:2605.15573v1 Announce Type: new Abstract: Multi-agent systems can solve complex tasks through collaboration between multiple Large Language Model agents. Existing collaboration frameworks typica

SGR: A Stepwise Reasoning Framework for LLMs with External Subgraph Generation

Model ReleasesDGX agent

arXiv:2605.16117v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated strong capabilities across diverse NLP applications, such as translation, text generation, and question a

SMMBench: A Benchmark for Source-Distributed Multimodal Agent Memory

Model ReleasesDGX agent

arXiv:2605.15710v1 Announce Type: new Abstract: Existing benchmarks for multimodal memory reasoning largely evaluate systems within pre-assembled contexts, but under-evaluate whether agents can use ev

Smoothie: Smoothing Diffusion on Token Embeddings for Text Generation

ResearchDGX agent

arXiv:2505.18853v2 Announce Type: replace Abstract: Diffusion models have achieved state-of-the-art performance in generating images, audio, and video, but their adaptation to text remains challenging

Stabilizing Knowledge, Promoting Reasoning: Dual-Token Constraints for RLVR

ResearchDGX agent

arXiv:2507.15778v2 Announce Type: replace Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has become an effective post-training method for improving the reasoning abilities of Large La

STS: Efficient Sparse Attention with Speculative Token Sparsity

Model ReleasesDGX agent

arXiv:2605.15508v1 Announce Type: cross Abstract: The quadratic complexity of attention imposes severe memory and computational bottlenecks on Large Language Model (LLM) inference. This challenge is p

Syntax Without Semantics: Teaching Large Language Models to Code in an Unseen Language

ResearchDGX agent

arXiv:2605.15607v1 Announce Type: new Abstract: Large language models (LLMs) achieve high pass rates on code generation benchmarks, yet whether they can transfer this ability to languages absent from

TemplateRL: Structured Template-Guided Reinforcement Learning for LLM Reasoning

SafetyDGX agent

arXiv:2505.15692v5 Announce Type: replace Abstract: Reinforcement learning (RL) has emerged as an effective paradigm for enhancing model reasoning. However, existing RL methods like GRPO typically rel

Toward LLMs Beyond English-Centric Development

ResearchDGX agent

arXiv:2605.15613v1 Announce Type: new Abstract: Through an analysis of sequences generated by open-weight large language models (LLMs), we demonstrate that LLMs are heavily biased toward English. Whil

VCG-Bench: Towards A Unified Visual-Centric Benchmark for Structured Generation and Editing

Model ReleasesDGX agent

arXiv:2605.15677v1 Announce Type: new Abstract: Despite the rapid advancements in Vision-Language Models (VLMs), a critical gap remains in their ability to handle structured, controllable diagrammatic

VSPO: Vector-Steered Policy Optimization for Behavioral Control

SafetyDGX agent

arXiv:2605.15604v1 Announce Type: cross Abstract: Modern language models often need to optimize a primary accuracy objective while also accommodating secondary behavioral preferences, such as verbosit

When Importance Sampling Misallocates Credit: Asymmetric Ratios for Outcome-Supervised RL

SafetyDGX agent

arXiv:2510.06062v2 Announce Type: replace Abstract: Reinforcement learning (RL) has shown great promise in large language models (LLMs) post-training, which typically rely on token-level clipping to m

When Latent Geometry Is Not Enough: Draft-Conditioned Latent Refinement for Non-Autoregressive Text Generation

SafetyDGX agent

arXiv:2605.15557v1 Announce Type: new Abstract: Continuous diffusion and flow models are attractive for non-autoregressive text generation because they can update all positions in parallel. A major di

Why are language models less surprised than humans? Testing the Parse Multiplicity Mismatch Hypothesis

ResearchDGX agent

arXiv:2605.15440v1 Announce Type: new Abstract: Surprisal theory posits that the processing difficulty of a word is determined by its predictability in context, offering a potential link between human

15 May 2026

A Calculus-Based Framework for Determining Vocabulary Size in End-to-End ASR

Model ReleasesDGX agent

arXiv:2605.14427v1 Announce Type: new Abstract: In hybrid automatic speech recognition (ASR) systems, the vocabulary size is unambiguous, typically determined by the number of phones, bi-phones, or tr

A Formative Study of Brief Affective Text as a Complement to Wearable Sensing for Longitudinal Student Health Monitoring

ResearchDGX agent

arXiv:2605.14360v1 Announce Type: cross Abstract: Wearable devices capture physiological and behavioral data with increasing fidelity, but the psychological context shaping these outcomes is difficult

A Hormone-inspired Emotion Layer for Transformer language models (HELT)

ResearchDGX agent

arXiv:2605.13858v1 Announce Type: cross Abstract: Large Language Models have demonstrated remarkable capabilities in generating contextually relevant and grammatically correct text. However, they fund

A Large Language Model Based Pipeline for Review of Systems Entity Recognition from Clinical Notes

Model ReleasesDGX agent

arXiv:2506.11067v3 Announce Type: replace Abstract: Objective: Develop a cost-effective, large language model (LLM)-based pipeline for automatically extracting Review of Systems (ROS) entities from cl

Auditing Agent Harness Safety

Model ReleasesDGX agent

arXiv:2605.14271v1 Announce Type: new Abstract: LLM agents increasingly run inside execution harnesses that dispatch tools, allocate resources, and route messages between specialized components. Howev

Automated Construction of a Knowledge Graph of Nuclear Fusion Energy for Effective Elicitation and Retrieval of Information

Model ReleasesDGX agent

arXiv:2504.07738v3 Announce Type: replace Abstract: In this document, we discuss a multi-step approach to automated construction of a knowledge graph, for structuring and representing domain-specific

Beyond Cosine Similarity: Zero-Initialized Residual Complex Projection for Aspect-Based Sentiment Analysis

ResearchDGX agent

arXiv:2603.28205v2 Announce Type: replace Abstract: Aspect-Based Sentiment Analysis (ABSA) faces critical challenges due to representation entanglement and false-negative collisions in real-valued emb

Beyond Mode-Seeking RL: Trajectory-Balance Post-Training for Diffusion Language Models

Model ReleasesDGX agent

arXiv:2605.13935v1 Announce Type: cross Abstract: Diffusion language models are a promising alternative to autoregressive models, yet post-training methods for them largely adapt reward-maximizing obj

BOOKMARKS: Efficient Active Storyline Memory for Role-playing

ResearchDGX agent

arXiv:2605.14169v1 Announce Type: new Abstract: Memory systems are critical for role-playing agents (RPAs) to maintain long-horizon consistency. However, existing RPA memory methods (e.g., profiling)

Chain-of-Procedure: Hierarchical Visual-Language Reasoning for Procedural QA

Model ReleasesDGX agent

arXiv:2605.14928v1 Announce Type: new Abstract: Recent advances in vision-language models (VLMs) have achieved impressive results on standard image-text tasks, yet their potential for visual procedure

Comparing Developer and LLM Biases in Code Evaluation

SafetyDGX agent

arXiv:2603.24586v2 Announce Type: replace-cross Abstract: As LLMs are increasingly used as judges in code applications, they should be evaluated in realistic interactive settings that capture partial

Confidence Estimation for LLMs in Multi-turn Interactions

ResearchDGX agent

arXiv:2601.02179v2 Announce Type: replace Abstract: While confidence estimation is a promising direction for mitigating hallucinations in Large Language Models (LLMs), current research overwhelmingly

Conversion of Lexicon-Grammar tables to LMF. Application to French

ResearchDGX agent

arXiv:2605.14816v1 Announce Type: new Abstract: We describe the first experiment of conversion of Lexicon-Grammar tables for French verbs into the Lexical Markup Framework (LMF) format. The Lexicon-Gr

CounselBench: A Large-Scale Expert Evaluation and Adversarial Benchmarking of Large Language Models in Mental Health Question Answering

Model ReleasesDGX agent

arXiv:2506.08584v4 Announce Type: replace Abstract: Medical question answering (QA) benchmarks often focus on multiple-choice or fact-based tasks, leaving open-ended answers to real patient questions

← Previous
1…7273747576…129
Next →