AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
19 May 2026

FishBack: Pullback Fisher Geometry for Optimal Activation Steering in Transformers

Local AiDGX agent

arXiv:2605.17231v1 Announce Type: cross Abstract: Activation steering methods modify intermediate representations of language models to control output behavior, but universally assume the activation s

Fix the Structural Bottleneck: Context Compression via Explicit Information Transmission

Model ReleasesDGX agent

arXiv:2602.03784v2 Announce Type: replace Abstract: Long-context LLM agents often struggle with growing token, memory, and latency costs, making efficient context compression essential for practical d

FOL2NS: Generating Natural Sentences from First-Order Logic

ResearchDGX agent

arXiv:2605.18155v1 Announce Type: new Abstract: Translating formal language into natural language is a foundational challenge in NLP, driving various downstream applications in semantic parsing, theor


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Forecasting Downstream Performance of LLMs With Proxy Metrics

ResearchDGX agent

arXiv:2605.18607v1 Announce Type: new Abstract: Progress in language model development is often driven by comparative decisions: which architecture to adopt, which pretraining corpus to use, or which

From BERT to T5: A Study of Named Entity Recognition

ResearchDGX agent

arXiv:2605.18462v1 Announce Type: new Abstract: Named entity recognition (NER) has been one of the essential preliminary steps in modern NLP applications. This report focuses on implementing the NER t

From Documents to Segments: A Contextual Reformulation for Topic Assignment

ApplicationsDGX agent

arXiv:2605.17714v1 Announce Type: new Abstract: Traditional topic modeling assigns a single topic to each document. In practice, however, many real-world documents, such as product reviews or open-end

From Isolated Scoring to Collaborative Ranking: A Comparison-Native Framework for LLM-Based Paper Evaluation

ResearchDGX agent

arXiv:2603.17588v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are currently applied to scientific paper evaluation by assigning an absolute score to each paper independently.

Gated KalmaNet: A Fading Memory Layer Through Test-Time Ridge Regression

Model ReleasesDGX agent

arXiv:2511.21016v3 Announce Type: replace-cross Abstract: Linear State-Space Models (SSMs) offer an efficient alternative to softmax Attention with constant memory and linear compute, but their lossy,

General Preference Reinforcement Learning

Model ReleasesDGX agent

arXiv:2605.18721v1 Announce Type: cross Abstract: Post-training has split large language model (LLM) alignment into two largely disconnected tracks. Online reinforcement learning (RL) with verifiable

Generative AI Advertising as a Problem of Trustworthy Commercial Intervention

AgentsDGX agent

arXiv:2605.18673v1 Announce Type: cross Abstract: Major deployed generative AI advertising systems preserve a visible boundary between commercial content and AI-generated responses. Yet empirical rese

Generative Artificial Intelligence for Literature Reviews

Model ReleasesDGX agent

arXiv:2605.16475v1 Announce Type: cross Abstract: Generative artificial intelligence (GenAI), based on large-language models (LLMs), such as ChatGPT, has taken organizations, academia, and the public

GUT-IS: A Data-Driven Approach to Integrating Constructs and Their Relations in Information Systems

ResearchDGX agent

arXiv:2605.18567v1 Announce Type: new Abstract: Structural equation modeling is widely used in IS research. However, inconsistent construct definitions impede the cumulative development of knowledge.

HalluScore: Large Language Model Hallucination Question Answering Benchmark

Model ReleasesDGX agent

arXiv:2605.17007v1 Announce Type: new Abstract: Large language models (LLMs) have achieved remarkable progress in natural language generation, but remain susceptible to hallucination. In response to g

HEED: Density-Weighted Residual Alignment for Hybrid Vision-Language Model Distillation

Model ReleasesDGX agent

arXiv:2605.17093v1 Announce Type: cross Abstract: Distilling vision-language models into faster hybrid architectures, such as 3:1 Mamba-2/attention mixes, is now standard practice for making inference

Helpful to a Fault: Measuring Illicit Assistance in Multi-Turn, Multilingual LLM Agents

SafetyDGX agent

arXiv:2602.16346v3 Announce Type: replace Abstract: LLM-based agents execute real-world workflows via tools and memory. These affordances enable ill-intended adversaries to also use these agents to ca

How Good LLMs Are at Answering Bangla Medical Visual Questions? Dataset and Benchmarking

Model ReleasesDGX agent

arXiv:2605.18111v1 Announce Type: new Abstract: Recent advancements in Large Language Models (LLMs) and Large Vision Language Models (LVLMs) have enabled general-purpose systems to demonstrate promisi

How Loud Rumbles Hit Newsstands: A Data Analysis of Coverage and Spatial Bias in German News about Landslides Around the World

SafetyDGX agent

arXiv:2605.18105v1 Announce Type: new Abstract: Landslides often hit newsstands due to their destructive and potentially fatal effects. News are a valuable source of information for creating or enrich

How Off-Policy Can GRPO Be? Mu-GRPO for Efficient LLM Reinforcement Learning

SafetyDGX agent

arXiv:2605.17570v1 Announce Type: cross Abstract: Group Relative Policy Optimization (GRPO) has been a key driver of recent progress in reinforcement learning with verifiable rewards (RLVR) for large

Hybrid Feature Combinations with CNN for Bangla Fake News Classification

ResearchDGX agent

arXiv:2605.17481v1 Announce Type: new Abstract: Nowadays, people in Bangladesh frequently rely on the internet and social media for daily news instead of traditional newspapers. However, the spread of

HyDRA: Hybrid Dynamic Routing Architecture for Heterogeneous LLM Pools

Model ReleasesDGX agent

arXiv:2605.17106v1 Announce Type: new Abstract: Production LLM deployments increasingly maintain heterogeneous model pools spanning order-of-magnitude cost differences. Existing routers make binary st

Implicit Hierarchical GRPO: Decoupling Tool Invocation from Execution for Tool-Integrated Mathematical Reasoning

SafetyDGX agent

arXiv:2605.18500v1 Announce Type: new Abstract: Large language models (LLMs) have increasingly leveraged tool invocation to enhance their reasoning capabilities. However, existing approaches typically

Infini-News: Efficiently Queryable Access to 1.3 Billion Processed Common Crawl News Articles

ResearchDGX agent

arXiv:2605.18337v1 Announce Type: new Abstract: Large-scale news corpora support a wide range of research in Computational Social Science and NLP, yet access remains constrained: commercial archives i

Information-Theoretic Storage Cost in Sentence Comprehension

ResearchDGX agent

arXiv:2602.18217v2 Announce Type: replace Abstract: Real-time sentence comprehension imposes a significant load on working memory, as comprehenders must maintain contextual information to anticipate f

Internalizing Tool Knowledge in Small Language Models via QLoRA Fine-Tuning

Model ReleasesDGX agent

arXiv:2605.17774v1 Announce Type: new Abstract: Large language models are increasingly used as planning components in agentic systems, but current tool-use pipelines often require full tool schemas to

iPOE: Interpretable Prompt Optimization via Explanations

TutorialsDGX agent

arXiv:2605.18113v1 Announce Type: new Abstract: Prompt optimization has often been framed as a discrete search problem to find high-performing and robust instructions for an LLM. However, the search r

JSPG: Dynamic Dictionary Filtering via Joint Semantic-Pinyin-Glyph Retrieval for Chinese Contextual ASR

ResearchDGX agent

arXiv:2605.16896v1 Announce Type: new Abstract: Contextual Automatic Speech Recognition (ASR) faces challenges with large-scale keyword dictionaries, as excessive irrelevant candidates introduce noise

Knowledge-to-Verification: Exploring RLVR for LLMs in Knowledge-Intensive Domains

ResearchDGX agent

arXiv:2605.18261v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has demonstrated promising potential to enhance the reasoning capabilities of large language model

KVDrive: A Holistic Multi-Tier KV Cache Management System for Long-Context LLM Inference

HardwareDGX agent

arXiv:2605.18071v1 Announce Type: new Abstract: Supporting long-context LLMs is challenging due to the substantial memory demands of the key-value (KV) cache. Existing offloading systems store the ful

Language Acquisition Device in Large Language Models

ResearchDGX agent

arXiv:2605.16758v1 Announce Type: new Abstract: Large Language Models (LLMs) remain substantially less data-efficient than humans. Pre-pretraining (PPT) on synthetic languages has been proposed to clo

Language-Switching Triggers Take a Latent Detour Through Language Models

Model ReleasesDGX agent

arXiv:2605.18646v1 Announce Type: new Abstract: Backdoor attacks on language models pose a growing security concern, yet the internal mechanisms by which a trigger sequence hijacks model computations

LaPA^2: Length-Aware Prefix and Prompt Attention Augmentation for Long-Form Controllable Text Generation

Model ReleasesDGX agent

arXiv:2508.04047v2 Announce Type: replace Abstract: Prefix-based methods have emerged as a promising paradigm for Controllable Text Generation (CTG) due to their parameter efficiency. However, while e

Large Language Models and Impossible Language Acquisition: 'False Promise' or an Overturn of our Current Perspective towards AI

ResearchDGX agent

arXiv:2602.08437v5 Announce Type: replace Abstract: In Chomsky's provocative critique 'The False Promise of CHATGPT,' Large Language Models (LLMs) are characterized as mere pattern predictors that do

Learning from Self-Debate: Preparing Reasoning Models for Multi-Agent Debate

AgentsDGX agent

arXiv:2601.22297v2 Announce Type: replace Abstract: The reasoning abilities of large language models (LLMs) have been substantially improved by reinforcement learning with verifiable rewards (RLVR). A

Learning to Reason without External Rewards

SafetyDGX agent

arXiv:2505.19590v5 Announce Type: replace-cross Abstract: Training large language models (LLMs) for complex reasoning via Reinforcement Learning with Verifiable Rewards (RLVR) is effective but limited

Learning Transferable Topology Priors for Multi-Agent LLM Collaboration Across Domains

SafetyDGX agent

arXiv:2605.17359v1 Announce Type: new Abstract: Large language model (LLM)-based multi-agent systems have shown strong potential for complex reasoning by coordinating specialized agents through struct

Leveraging Multimodal Self-Consistency Reasoning in Coding Motivational Interviewing for Alcohol Use Reduction

ResearchDGX agent

arXiv:2605.12987v2 Announce Type: replace Abstract: BACKGROUND: Coding Motivational Interviewing (MI) sessions is essential for understanding client behaviors and predicting outcomes, but it requires

Linguistic Uncertainty and Reply Engagement on X: A Cross-Domain Replication of the Uncertainty-Reply Asymmetry

SafetyDGX agent

arXiv:2605.16289v1 Announce Type: cross Abstract: Linguistic uncertainty is common in social media, but its relationship with engagement remains unclear across languages and topics. Using 2,258 Englis

LISTEN to Your Preferences: An LLM Framework for Multi-Objective Selection

SafetyDGX agent

arXiv:2510.25799v2 Announce Type: replace Abstract: Human experts often struggle to select the best option from a large set of items with multiple competing objectives, a process bottlenecked by the d

LLM Agents Are the Antidote to Walled Gardens

ApplicationsDGX agent

arXiv:2506.23978v3 Announce Type: replace-cross Abstract: While the Internet's core infrastructure was designed to be open and universal, today's application layer is dominated by closed, proprietary

LLM-Based Intelligent Notification Composition: From Static Personalization to Context-Aware Persuasive Messaging

TutorialsDGX agent

arXiv:2605.16264v1 Announce Type: cross Abstract: Push notifications remain among the most direct channels through which digital platforms engage users, yet existing approaches have invested heavily i

LLMs for automatic annotation of Mandarin narrative transcripts

Local AiDGX agent

arXiv:2605.17205v1 Announce Type: new Abstract: Linguistic annotation of transcribed speech is essential for research in language acquisition, language disorders, and sociolinguistics, yet remains lab

LLMs in Qualitative Research: Opportunities, Limitations, and Practical Considerations

Model ReleasesDGX agent

arXiv:2605.16538v1 Announce Type: cross Abstract: This paper examines the opportunities, limitations, and practical considerations associated with the use of large language models (LLMs) in qualitativ

MA^{2}P: A Meta-Cognitive Autonomous Intelligent Agents Framework for Complex Persuasion

AgentsDGX agent

arXiv:2605.18572v1 Announce Type: new Abstract: Persuasive dialogue generation plays a vital role in decision-making, negotiation, counseling, and behavior change, yet it remains a challenging problem

Medical Context Distorts Decisions in Clinical Vision Language Models

SafetyDGX agent

arXiv:2605.17436v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly proposed for clinical decision support, yet their reliability in real-world scenarios that require inte

MentalBench: A DSM-Grounded Benchmark for Evaluating Psychiatric Diagnostic Capability of Large Language Models

Model ReleasesDGX agent

arXiv:2602.12871v2 Announce Type: replace Abstract: Large language models (LLMs) have attracted growing interest as supportive tools for psychiatric assessment and clinical decision support. However,

Merlin's Whisper: Enabling Efficient Reasoning in Large Language Models via Black-box Persuasive Prompting

Model ReleasesDGX agent

arXiv:2510.10528v3 Announce Type: replace Abstract: Large reasoning models (LRMs) have demonstrated remarkable proficiency in tackling complex tasks through step-by-step thinking. However, this length

MiniGPT: Rebuilding GPT from First Principles

Model ReleasesDGX agent

arXiv:2605.17398v1 Announce Type: new Abstract: This paper presents MiniGPT, a compact from-scratch implementation of GPT-style autoregressive language modeling in PyTorch. The aim is to rebuild the c

MixSD: Mixed Contextual Self-Distillation for Knowledge Injection

Model ReleasesDGX agent

arXiv:2605.16865v1 Announce Type: new Abstract: Supervised fine-tuning (SFT) is widely used to inject new knowledge into language models, but it often degrades pretrained capabilities such as reasonin

Mixture of Experts for Low-Resource LLMs

Model ReleasesDGX agent

arXiv:2605.17598v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) architectures enable efficient model scaling, yet expert routing behavior across underrepresented languages remains poorly unde

Monitoring the Internal Monologue: Probe Trajectories Reveal Reasoning Dynamics

SafetyDGX agent

arXiv:2605.18549v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) introduce new opportunities for safety monitoring through their Chain of Thought (CoT) reasoning. However, CoT is not alwa

Multilingual and Multimodal LLMs in the Wild: Building for Low-Resource Languages

SafetyDGX agent

arXiv:2605.17152v1 Announce Type: new Abstract: Multimodal LLMs are evolving from vision-language to tri-modality that see, hear, and read, yet pipelines and benchmarks remain English-centric and comp

Multilingual OCR-Aware Fine-Tuning and Prompt-Guided Chain-of-Thought Reasoning for Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2605.16409v1 Announce Type: cross Abstract: Optical character recognition (OCR) and multilingual text understanding remain major failure modes of multimodal large language models (MLLMs), partic

NewsLens: A Multi-Agent Framework for Adversarial News Bias Navigation

Model ReleasesDGX agent

arXiv:2605.17364v1 Announce Type: new Abstract: Media bias detection has predominantly been framed as a classification task: assign a political label to an article or outlet. We argue this framing is

PaliBench: A Multi-Reference Blueprint for Classical Language Translation Benchmarks

Model ReleasesDGX agent

arXiv:2605.16881v1 Announce Type: new Abstract: Digital humanities projects increasingly rely on machine translation and large language models to widen access to classical, religious, and otherwise un

PEGRL: Improving Machine Translation by Post-Editing Guided Reinforcement Learning

Model ReleasesDGX agent

arXiv:2602.03352v2 Announce Type: replace Abstract: Reinforcement learning (RL) has shown strong promise for LLM-based machine translation, with recent methods such as GRPO demonstrating notable gains

PPAI: Enabling Personalized LLM Agent Interoperability for Collaborative Edge Intelligence

Local AiDGX agent

arXiv:2605.18067v1 Announce Type: new Abstract: Deploying large language model (LLM) on edge device enables personalized LLM agents for various users. The growing availability of diverse personalized

PQR: A Framework to Generate Diverse and Realistic User Queries that Elicit QA Agent Failures

SafetyDGX agent

arXiv:2605.16551v1 Announce Type: new Abstract: Evaluating LLM-based agents remains challenging because identifying meaningful failure cases often requires substantial human effort to design realistic

Presupposition and Reasoning in Conditionals: A Theory-Based Study of Humans and LLMs

SafetyDGX agent

arXiv:2605.18352v1 Announce Type: new Abstract: Presupposition projection in conditionals is central to theories of meaning and pragmatics, yet it remains largely unevaluated in large language models.

Proof-Carrying Certificates for LLM Pipelines: A Trust-Boundary Architecture

AgentsDGX agent

arXiv:2605.16407v1 Announce Type: cross Abstract: We present a framework for verifying the deterministic structured computations surrounding a large language model rather than the model itself, extend

Protection Is (Nearly) All You Need: Structural Protection Dominates Scoring in Globally Capped KV Eviction

Model ReleasesDGX agent

arXiv:2605.18053v1 Announce Type: cross Abstract: We study KV cache eviction under a shared globally capped decode-time harness. Seven policies (LRU, H2O, SnapKV, StreamingLLM, Ada-KV, QUEST, Random)

← Previous
1…7071727374…129
Next →