AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlog
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,646 results
Model Releases

VeriInteresting: An Empirical Study of Model Prompt Interactions in Verilog Code Generation

DGX agent

arXiv:2603.08715v2 Announce Type: replace-cross Abstract: Rapid advances in language models (LMs) have created new opportunities for automated code generation while complicating trade-offs between mod

model-releasesarxiv-cs-cl
14 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

VisText-Mosquito: A Unified Multimodal Dataset for Visual Detection, Segmentation, and Textual Explanation on Mosquito Breeding Sites

DGX agent

arXiv:2506.14629v3 Announce Type: replace-cross Abstract: Mosquito-borne diseases pose a major global health risk, requiring early detection and proactive control of breeding sites to prevent outbreak

researcharxiv-cs-cl
14 Apr 2026
Research

Visual Late Chunking: An Empirical Study of Contextual Chunking for Efficient Visual Document Retrieval

DGX agent

arXiv:2604.10167v1 Announce Type: cross Abstract: Multi-vector models dominate Visual Document Retrieval (VDR) due to their fine-grained matching capabilities, but their high storage and computational

researcharxiv-cs-cl
14 Apr 2026
Model Releases

VLN-NF: Feasibility-Aware Vision-and-Language Navigation with False-Premise Instructions

DGX agent

arXiv:2604.10533v1 Announce Type: cross Abstract: Conventional Vision-and-Language Navigation (VLN) benchmarks assume instructions are feasible and the referenced target exists, leaving agents ill-equ

model-releasesarxiv-cs-cl
14 Apr 2026
Safety

Weird Generalization is Weirdly Brittle

DGX agent

arXiv:2604.10022v1 Announce Type: new Abstract: Weird generalization is a phenomenon in which models fine-tuned on data from a narrow domain (e.g. insecure code) develop surprising traits that manifes

safetyarxiv-cs-cl
14 Apr 2026
Applications

What Do Vision-Language Models Encode for Personalized Image Aesthetics Assessment?

DGX agent

arXiv:2604.11374v1 Announce Type: cross Abstract: Personalized image aesthetics assessment (PIAA) is an important research problem with practical real-world applications. While methods based on vision

applicationsarxiv-cs-cl
14 Apr 2026
Safety

What Factors Affect LLMs and RLLMs in Financial Question Answering?

DGX agent

arXiv:2507.08339v4 Announce Type: replace Abstract: Recently, large language models (LLMs) and reasoning large language models (RLLMs) have gained considerable attention from many researchers. RLLMs e

safetyarxiv-cs-cl
14 Apr 2026
Model Releases

When Meaning Isn't Literal: Exploring Idiomatic Meaning Across Languages and Modalities

DGX agent

arXiv:2604.10787v1 Announce Type: new Abstract: Idiomatic reasoning, deeply intertwined with metaphor and culture, remains a blind spot for contemporary language models, whose progress skews toward su

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Who Wrote This Line? Evaluating the Detection of LLM-Generated Classical Chinese Poetry

DGX agent

arXiv:2604.10101v1 Announce Type: new Abstract: The rapid development of large language models (LLMs) has extended text generation tasks into the literary domain. However, AI-generated literary creati

model-releasesarxiv-cs-cl
14 Apr 2026
Local Ai

Why Code, Why Now: An Information-Theoretic Perspective on the Limits of Machine Learning

DGX agent

arXiv:2602.13934v4 Announce Type: replace-cross Abstract: This paper offers a new perspective on the limits of machine learning: the ceiling on progress is set not by model size or algorithm choice bu

local-aiarxiv-cs-cl
14 Apr 2026
Applications

Why Don't You Know? Evaluating the Impact of Uncertainty Sources on Uncertainty Quantification in LLMs

DGX agent

arXiv:2604.10495v1 Announce Type: new Abstract: As Large Language Models (LLMs) are increasingly deployed in real-world applications, reliable uncertainty quantification (UQ) becomes critical for safe

applicationsarxiv-cs-cl
14 Apr 2026
Model Releases

Why Supervised Fine-Tuning Fails to Learn: A Systematic Study of Incomplete Learning in Large Language Models

DGX agent

arXiv:2604.10079v1 Announce Type: new Abstract: Supervised Fine-Tuning (SFT) is the standard approach for adapting large language models (LLMs) to downstream tasks. However, we observe a persistent fa

model-releasesarxiv-cs-cl
14 Apr 2026
Safety

YIELD: A Large-Scale Dataset and Evaluation Framework for Information Elicitation Agents

DGX agent

arXiv:2604.10968v1 Announce Type: new Abstract: Most conversational agents (CAs) are designed to satisfy user needs through user-driven interactions. However, many real-world settings, such as academi

safetyarxiv-cs-cl
14 Apr 2026
Model Releases

ZARA: Training-Free Motion Time-Series Reasoning via Evidence-Grounded LLM Agents

DGX agent

arXiv:2508.04038v2 Announce Type: replace Abstract: Motion sensor time-series are central to Human Activity Recognition (HAR), yet conventional approaches are constrained to fixed activity sets and ty

model-releasesarxiv-cs-cl
14 Apr 2026
Safety

A Representation-Level Assessment of Bias Mitigation in Foundation Models

DGX agent

arXiv:2604.08561v1 Announce Type: new Abstract: We investigate how successful bias mitigation reshapes the embedding space of encoder-only and decoder-only foundation models, offering an internal audi

safetyarxiv-cs-cl
13 Apr 2026
Research

Across the Levels of Analysis: Explaining Predictive Processing in Humans Requires More Than Machine-Estimated Probabilities

DGX agent

arXiv:2604.09466v1 Announce Type: new Abstract: Under the lens of Marr's levels of analysis, we critique and extend two claims about language models (LMs) and language processing: first, that predicti

researcharxiv-cs-cl
13 Apr 2026
Model Releases

Agentic Jackal: Live Execution and Semantic Value Grounding for Text-to-JQL

DGX agent

arXiv:2604.09470v1 Announce Type: new Abstract: Translating natural language into Jira Query Language (JQL) requires resolving ambiguous field references, instance-specific categorical values, and com

model-releasesarxiv-cs-cl
13 Apr 2026
Model Releases

Anchored Sliding Window: Toward Robust and Imperceptible Linguistic Steganography

DGX agent

arXiv:2604.09066v1 Announce Type: new Abstract: Linguistic steganography based on language models typically assumes that steganographic texts are transmitted without alteration, making them fragile to

model-releasesarxiv-cs-cl
13 Apr 2026
Research

Arbitration Failure, Not Perceptual Blindness: How Vision-Language Models Resolve Visual-Linguistic Conflicts

DGX agent

arXiv:2604.09364v1 Announce Type: cross Abstract: When a Vision-Language Model (VLM) sees a blue banana and answers 'yellow', is the problem of perception or arbitration? We explore the question in te

researcharxiv-cs-cl
13 Apr 2026
Research

Attention-Based Sampler for Diffusion Language Models

DGX agent

arXiv:2604.08564v1 Announce Type: new Abstract: Auto-regressive models (ARMs) have established a dominant paradigm in language modeling. However, their strictly sequential decoding paradigm imposes fu

researcharxiv-cs-cl
13 Apr 2026
Model Releases

Automated Instruction Revision (AIR): A Structured Comparison of Task Adaptation Strategies for LLM

DGX agent

arXiv:2604.09418v1 Announce Type: new Abstract: This paper studies Automated Instruction Revision (AIR), a rule-induction-based method for adapting large language models (LLMs) to downstream tasks usi

model-releasesarxiv-cs-cl
13 Apr 2026
Model Releases

BEDTime: A Unified Benchmark for Automatically Describing Time Series

DGX agent

arXiv:2509.05215v3 Announce Type: replace Abstract: Recent works propose complex multi-modal models that handle both time series and language, ultimately claiming high performance on complex tasks lik

model-releasesarxiv-cs-cl
13 Apr 2026
Model Releases

Breaking Block Boundaries: Anchor-based History-stable Decoding for Diffusion Large Language Models

DGX agent

arXiv:2604.08964v1 Announce Type: new Abstract: Diffusion Large Language Models (dLLMs) have recently become a promising alternative to autoregressive large language models (ARMs). Semi-autoregressive

model-releasesarxiv-cs-cl
13 Apr 2026
Agents

CodeScout: Contextual Problem Statement Enhancement for Software Agents

DGX agent

arXiv:2603.05744v2 Announce Type: replace Abstract: Current AI-powered code assistance tools often struggle with poorly-defined problem statements that lack sufficient task context and requirements sp

agentsarxiv-cs-cl
13 Apr 2026
Applications

Confident in a Confidence Score: Investigating the Sensitivity of Confidence Scores to Supervised Fine-Tuning

DGX agent

arXiv:2604.08974v1 Announce Type: new Abstract: Uncertainty quantification is a set of techniques that measure confidence in language models. They can be used, for example, to detect hallucinations or

applicationsarxiv-cs-cl
13 Apr 2026
Model Releases

Cross-Lingual Attention Distillation with Personality-Informed Generative Augmentation for Multilingual Personality Recognition

DGX agent

arXiv:2604.08851v1 Announce Type: new Abstract: While significant work has been done on personality recognition, the lack of multilingual datasets remains an unresolved challenge. To address this, we

model-releasesarxiv-cs-cl
13 Apr 2026
Safety

EthicMind: A Risk-Aware Framework for Ethical-Emotional Alignment in Multi-Turn Dialogue

DGX agent

arXiv:2604.09265v1 Announce Type: new Abstract: Intelligent dialogue systems are increasingly deployed in emotionally and ethically sensitive settings, where failures in either emotional attunement or

safetyarxiv-cs-cl
13 Apr 2026
Research

EVOKE: Emotion Vocabulary Of Korean and English

DGX agent

arXiv:2602.10414v2 Announce Type: replace Abstract: This paper introduces EVOKE (Emotion Vocabulary of Korean and English), a Korean-English parallel dataset of emotion words. The dataset offers compr

researcharxiv-cs-cl
13 Apr 2026
Model Releases

EXAONE 4.5 Technical Report

DGX agent

arXiv:2604.08644v1 Announce Type: new Abstract: This technical report introduces EXAONE 4.5, the first open-weight vision language model released by LG AI Research. EXAONE 4.5 is architected by integr

model-releasesarxiv-cs-cl
13 Apr 2026
Research

Exploiting Web Search Tools of AI Agents for Data Exfiltration

DGX agent

arXiv:2510.09093v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are now routinely used to autonomously execute complex tasks, from natural language processing to dynamic workflo

researcharxiv-cs-cl
13 Apr 2026
Research

Exploring Cross-lingual Latent Transplantation: Mutual Opportunities and Open Challenges

DGX agent

arXiv:2412.12686v3 Announce Type: replace Abstract: Current large language models (LLMs) often exhibit imbalances in multilingual capabilities and cultural adaptability, largely attributed to their En

researcharxiv-cs-cl
13 Apr 2026
Model Releases

Facet-Level Tracing of Evidence Uncertainty and Hallucination in RAG

DGX agent

arXiv:2604.09174v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) aims to reduce hallucination by grounding answers in retrieved evidence, yet hallucinated answers remain common eve

model-releasesarxiv-cs-cl
13 Apr 2026
Agents

Fast-dVLM: Efficient Block-Diffusion VLM via Direct Conversion from Autoregressive VLM

DGX agent

arXiv:2604.06832v2 Announce Type: replace Abstract: Vision-language models (VLMs) predominantly rely on autoregressive decoding, which generates tokens one at a time and fundamentally limits inference

agentsarxiv-cs-cl
13 Apr 2026
Research

Few-Shot Contrastive Adaptation for Audio Abuse Detection in Low-Resource Indic Languages

DGX agent

arXiv:2604.09094v1 Announce Type: cross Abstract: Abusive speech detection is becoming increasingly important as social media shifts towards voice-based interaction, particularly in multilingual and l

researcharxiv-cs-cl
13 Apr 2026
Safety

FP8-RL: A Practical and Stable Low-Precision Stack for LLM Reinforcement Learning

DGX agent

arXiv:2601.18150v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) for large language models (LLMs) is increasingly bottlenecked by rollout (generation), where long output sequence

safetyarxiv-cs-cl
13 Apr 2026
Model Releases

From Reasoning to Agentic: Credit Assignment in Reinforcement Learning for Large Language Models

DGX agent

arXiv:2604.09459v1 Announce Type: new Abstract: Reinforcement learning (RL) for large language models (LLMs) increasingly relies on sparse, outcome-level rewards -- yet determining which actions withi

model-releasesarxiv-cs-cl
13 Apr 2026
Research

Grammar as a Behavioral Biometric: Using Cognitively Motivated Grammar Models for Authorship Verification

DGX agent

arXiv:2403.08462v3 Announce Type: replace Abstract: Authorship Verification (AV) is a key area of research in digital text forensics, which addresses the fundamental question of whether two texts were

researcharxiv-cs-cl
13 Apr 2026
Model Releases

GRASP: Grounded CoT Reasoning with Dual-Stage Optimization for Multimodal Sarcasm Target Identification

DGX agent

arXiv:2604.08879v1 Announce Type: new Abstract: Moving beyond the traditional binary classification paradigm of Multimodal Sarcasm Detection, Multimodal Sarcasm Target Identification (MSTI) presents a

model-releasesarxiv-cs-cl
13 Apr 2026
Research

Growing a Multi-head Twig via Distillation and Reinforcement Learning to Accelerate Large Vision-Language Models

DGX agent

arXiv:2503.14075v3 Announce Type: replace-cross Abstract: Large vision-language models (VLMs) have demonstrated remarkable capabilities in open-world multimodal understanding, yet their high computati

researcharxiv-cs-cl
13 Apr 2026
Safety

Hierarchical Alignment: Enforcing Hierarchical Instruction-Following in LLMs through Logical Consistency

DGX agent

arXiv:2604.09075v1 Announce Type: new Abstract: Large language models increasingly operate under multiple instructions from heterogeneous sources with different authority levels, including system poli

safetyarxiv-cs-cl
13 Apr 2026
Research

Is More Data Worth the Cost? Dataset Scaling Laws in a Tiny Attention-Only Decoder

DGX agent

arXiv:2604.09389v1 Announce Type: cross Abstract: Training Transformer language models is expensive, as performance typically improves with increasing dataset size and computational budget. Although s

researcharxiv-cs-cl
13 Apr 2026
Local Ai

LADR: Locality-Aware Dynamic Rescue for Efficient Text-to-Image Generation with Diffusion Large Language Models

DGX agent

arXiv:2603.13450v2 Announce Type: replace-cross Abstract: Discrete Diffusion Language Models have emerged as a compelling paradigm for unified multimodal generation, yet their deployment is hindered b

local-aiarxiv-cs-cl
13 Apr 2026
Research

Localizing Task Recognition and Task Learning in In-Context Learning via Attention Head Analysis

DGX agent

arXiv:2509.24164v2 Announce Type: replace Abstract: We investigate the mechanistic underpinnings of in-context learning (ICL) in large language models by reconciling two dominant perspectives: the com

researcharxiv-cs-cl
13 Apr 2026
Safety

MAB-DQA: Addressing Query Aspect Importance in Document Question Answering with Multi-Armed Bandits

DGX agent

arXiv:2604.08952v1 Announce Type: new Abstract: Document Question Answering (DQA) involves generating answers from a document based on a user's query, representing a key task in document understanding

safetyarxiv-cs-cl
13 Apr 2026
Model Releases

Many Ways to Be Fake: Benchmarking Fake News Detection Under Strategy-Driven AI Generation

DGX agent

arXiv:2604.09514v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have enabled the large-scale generation of highly fluent and deceptive news-like content. While prior wo

model-releasesarxiv-cs-cl
13 Apr 2026
Model Releases

MedConceal: A Benchmark for Clinical Hidden-Concern Reasoning Under Partial Observability

DGX agent

arXiv:2604.08788v1 Announce Type: new Abstract: Patient-clinician communication is an asymmetric-information problem: patients often do not disclose fears, misconceptions, or practical barriers unless

model-releasesarxiv-cs-cl
13 Apr 2026
Model Releases

Mnemis: Dual-Route Retrieval on Hierarchical Graphs for Long-Term LLM Memory

DGX agent

arXiv:2602.15313v2 Announce Type: replace Abstract: AI Memory, specifically how models organizes and retrieves historical messages, becomes increasingly valuable to Large Language Models (LLMs), yet e

model-releasesarxiv-cs-cl
13 Apr 2026
Safety

MSMO-ABSA: Multi-Scale and Multi-Objective Optimization for Cross-Lingual Aspect-Based Sentiment Analysis

DGX agent

arXiv:2502.13718v2 Announce Type: replace Abstract: Aspect-based sentiment analysis (ABSA) garnered growing research interest in multilingual contexts in the past. However, the majority of the studies

safetyarxiv-cs-cl
13 Apr 2026
← Previous
1…153154155156157…160
Next →