AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Model Releases

Emergent Misalignment via In-Context Learning: Narrow in-context examples can produce broadly misaligned LLMs

DGX agent

arXiv:2510.11288v4 Announce Type: replace Abstract: Recent work has shown that narrow finetuning can produce broadly misaligned LLMs, a phenomenon termed emergent misalignment (EM). While concerning,

model-releasesarxiv-cs-cl
21 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

Emergent Structured Representations Support Flexible In-Context Inference in Large Language Models

DGX agent

arXiv:2602.07794v3 Announce Type: replace Abstract: Large language models (LLMs) exhibit emergent behaviors suggestive of human-like reasoning. While recent work has identified structured conceptual r

researcharxiv-cs-cl
21 Apr 2026
Research

emg2speech: Synthesizing speech from electromyography using self-supervised speech models

DGX agent

arXiv:2510.23969v2 Announce Type: replace-cross Abstract: We present a neuromuscular speech interface that translates electromyographic (EMG) signals recorded from orofacial muscles during speech arti

researcharxiv-cs-cl
21 Apr 2026
Research

Emotion Collider: Dual Hyperbolic Mirror Manifolds for Sentiment Recovery via Anti Emotion Reflection

DGX agent

arXiv:2602.16161v3 Announce Type: replace-cross Abstract: Emotional expression underpins natural communication and effective human-computer interaction. We present Emotion Collider (EC-Net), a hyperbo

researcharxiv-cs-cl
21 Apr 2026
Model Releases

Employing General-Purpose and Biomedical Large Language Models with Advanced Prompt Engineering for Pharmacoepidemiologic Study Design

DGX agent

arXiv:2604.17988v1 Announce Type: new Abstract: Background: The potential of large language models (LLMs) to automate and support pharmacoepidemiologic study design is an emerging area of interest, ye

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

Empowering Multi-Turn Tool-Integrated Agentic Reasoning with Group Turn Policy Optimization

DGX agent

arXiv:2511.14846v2 Announce Type: replace-cross Abstract: Training Large Language Models (LLMs) for multi-turn Tool-Integrated Reasoning (TIR) - where models iteratively reason, generate code, and ver

safetyarxiv-cs-cl
21 Apr 2026
Hardware

Enabling AI ASICs for Zero Knowledge Proof

DGX agent

arXiv:2604.17808v1 Announce Type: cross Abstract: Zero-knowledge proof (ZKP) provers remain costly because multi-scalar multiplication (MSM) and number-theoretic transforms (NTTs) dominate runtime as

hardwarearxiv-cs-cl
21 Apr 2026
Research

Enabling Stroke-Level Structural Analysis of Hieroglyphic Scripts without Language-Specific Priors

DGX agent

arXiv:2601.05508v2 Announce Type: replace-cross Abstract: Hieroglyphs, as logographic writing systems, encode rich semantic and cultural information within their internal structural composition. Yet,

researcharxiv-cs-cl
21 Apr 2026
Model Releases

End-to-end Listen, Look, Speak and Act

DGX agent

arXiv:2510.16756v2 Announce Type: replace-cross Abstract: Human interaction is inherently multimodal and full-duplex: we listen while watching, speak while acting, and fluidly adapt to turn-taking and

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Enhancing Trust in Large Language Models via Uncertainty-Calibrated Fine-Tuning

DGX agent

arXiv:2412.02904v2 Announce Type: replace Abstract: Large language models (LLMs) have revolutionized the field of natural language processing with their impressive reasoning and question-answering cap

researcharxiv-cs-cl
21 Apr 2026
Research

Erase to Improve: Erasable Reinforcement Learning for Search-Augmented LLMs

DGX agent

arXiv:2510.00861v2 Announce Type: replace Abstract: While search-augmented large language models (LLMs) exhibit impressive capabilities, their reliability in complex multi-hop reasoning remains limite

researcharxiv-cs-cl
21 Apr 2026
Model Releases

ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection

DGX agent

arXiv:2410.04509v3 Announce Type: replace Abstract: As the field of Multimodal Large Language Models (MLLMs) continues to evolve, their potential to revolutionize artificial intelligence is particular

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

ESsEN: Training Compact Discriminative Vision-Language Transformers in a Low-Resource Setting

DGX agent

arXiv:2604.18452v1 Announce Type: cross Abstract: Vision-language modeling is rapidly increasing in popularity with an ever expanding list of available models. In most cases, these vision-language mod

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Establishing a Scale for Kullback-Leibler Divergence in Language Models Across Various Settings

DGX agent

arXiv:2505.15353v3 Announce Type: replace Abstract: Log-likelihood vectors define a common space for comparing language models as probability distributions, enabling unified comparisons across heterog

researcharxiv-cs-cl
21 Apr 2026
Research

Estimating Commonsense Plausibility through Semantic Shifts

DGX agent

arXiv:2502.13464v2 Announce Type: replace Abstract: Commonsense plausibility estimation is critical for evaluating language models (LMs), yet existing generative approaches--reliant on likelihoods or

researcharxiv-cs-cl
21 Apr 2026
Research

Evalet: Evaluating Large Language Models through Functional Fragmentation

DGX agent

arXiv:2509.11206v4 Announce Type: replace-cross Abstract: Practitioners increasingly rely on Large Language Models (LLMs) to evaluate generative AI outputs through 'LLM-as-a-Judge' approaches. However

researcharxiv-cs-cl
21 Apr 2026
Tutorials

Evaluating Adaptive Personalization of Educational Readings with Simulated Learners

DGX agent

arXiv:2604.16744v1 Announce Type: new Abstract: We present a framework for evaluating adaptive personalization of educational reading materials with theory-grounded simulated learners. The system buil

tutorialsarxiv-cs-cl
21 Apr 2026
Research

Evaluating the Impact of Verbal Multiword Expressions on Machine Translation

DGX agent

arXiv:2508.17458v2 Announce Type: replace Abstract: Verbal multiword expressions (VMWEs) remain difficult for machine translation because their meanings are often not recoverable from their component

researcharxiv-cs-cl
21 Apr 2026
Model Releases

Evaluating Tool-Using Language Agents: Judge Reliability, Propagation Cascades, and Runtime Mitigation in AgentProp-Bench

DGX agent

arXiv:2604.16706v1 Announce Type: cross Abstract: Automated evaluation of tool-using large language model (LLM) agents is widely assumed to be reliable, but this assumption has rarely been validated a

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

Evidence-Augmented Policy Optimization with Reward Co-Evolution for Long-Context Reasoning

DGX agent

arXiv:2601.10306v2 Announce Type: replace-cross Abstract: While Reinforcement Learning (RL) has advanced LLM reasoning, applying it to long-context scenarios is hindered by sparsity of outcome rewards

safetyarxiv-cs-cl
21 Apr 2026
Safety

Evolutionary Negative Module Pruning for Better LoRA Merging

DGX agent

arXiv:2604.17753v1 Announce Type: cross Abstract: Merging multiple Low-Rank Adaptation (LoRA) experts into a single backbone is a promising approach for efficient multi-task deployment. While existing

safetyarxiv-cs-cl
21 Apr 2026
Safety

Explanation Bias is a Product: Revealing the Hidden Lexical and Position Preferences in Post-Hoc Feature Attribution

DGX agent

arXiv:2512.11108v3 Announce Type: replace Abstract: Good quality explanations strengthen the understanding of language models and data. Feature attribution methods, such as Integrated Gradient, are a

safetyarxiv-cs-cl
21 Apr 2026
Research

Exploring Concreteness Through a Figurative Lens

DGX agent

arXiv:2604.18296v1 Announce Type: new Abstract: Static concreteness ratings are widely used in NLP, yet a word's concreteness can shift with context, especially in figurative language such as metaphor

researcharxiv-cs-cl
21 Apr 2026
Research

Expressing Social Emotions: Misalignment Between LLMs and Human Cultural Emotion Norms

DGX agent

arXiv:2604.16757v1 Announce Type: new Abstract: The expression of emotions that serve social purposes, such as asserting independence or fostering interdependence, is central to human interactions and

researcharxiv-cs-cl
21 Apr 2026
Safety

Faithfulness vs. Safety: Evaluating LLM Behavior Under Counterfactual Medical Evidence

DGX agent

arXiv:2601.11886v2 Announce Type: replace Abstract: In high-stakes domains like medicine, it may be generally desirable for models to faithfully adhere to the context provided. But what happens if the

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

FaithLens: Detecting and Explaining Faithfulness Hallucination

DGX agent

arXiv:2512.20182v3 Announce Type: replace Abstract: Recognizing whether outputs from large language models (LLMs) contain faithfulness hallucination is crucial for real-world applications, e.g., retri

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Finding Culture-Sensitive Neurons in Vision-Language Models

DGX agent

arXiv:2510.24942v2 Announce Type: replace-cross Abstract: Despite their impressive performance, vision-language models (VLMs) still struggle on culturally situated inputs. To understand how VLMs proce

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

FLARE: Task-agnostic embedding model evaluation through a normalization process

DGX agent

arXiv:2604.17344v1 Announce Type: cross Abstract: When task-specific labels are not available, it becomes difficult to select an embedding model for a specific target corpus. Existing labelless measur

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

FLiP: Towards understanding and interpreting multimodal multilingual sentence embeddings

DGX agent

arXiv:2604.18109v1 Announce Type: new Abstract: This paper presents factorized linear projection (FLiP) models for understanding pretrained sentence embedding spaces. We train FLiP models to recover t

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Follow the Path: Reasoning over Knowledge Graph Paths to Improve Large Language Model Factuality

DGX agent

arXiv:2505.11140v3 Announce Type: replace Abstract: We introduce fs1, a simple yet effective method that improves the factuality of reasoning traces by collecting them from large reasoning models and

researcharxiv-cs-cl
21 Apr 2026
Local Ai

Forest Before Trees: Latent Superposition for Efficient Visual Reasoning

DGX agent

arXiv:2601.06803v2 Announce Type: replace Abstract: While Chain-of-Thought empowers Large Vision-Language Models with multi-step reasoning, explicit textual rationales suffer from an information bandw

local-aiarxiv-cs-cl
21 Apr 2026
Model Releases

FOREVER: Forgetting Curve-Inspired Memory Replay for Language Model Continual Learning

DGX agent

arXiv:2601.03938v2 Announce Type: replace-cross Abstract: Continual learning (CL) for large language models (LLMs) aims to enable sequential knowledge acquisition without catastrophic forgetting. Memo

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Forget What Matters, Keep the Rest: Selective Unlearning of Informative Tokens

DGX agent

arXiv:2604.17785v1 Announce Type: new Abstract: Unlearning in large language models (LLMs) has emerged as a promising safeguard against adversarial behaviors. When the forgetting loss is applied unifo

researcharxiv-cs-cl
21 Apr 2026
Research

Foundational Study on Authorship Attribution of Japanese Web Reviews for Actor Analysis

DGX agent

arXiv:2604.16376v1 Announce Type: new Abstract: This study investigates the applicability of authorship attribution based on stylistic features to support actor analysis in threat intelligence. As a f

researcharxiv-cs-cl
21 Apr 2026
Model Releases

Frankentext: Stitching random text fragments into long-form narratives

DGX agent

arXiv:2505.18128v4 Announce Type: replace Abstract: We introduce Frankentexts, a long-form narrative generation paradigm that treats an LLM as a composer of existing texts rather than as an author. Gi

model-releasesarxiv-cs-cl
21 Apr 2026
Research

FreezeEmpath: Efficient Training for Empathetic Spoken Chatbots with Frozen LLMs

DGX agent

arXiv:2604.18159v1 Announce Type: new Abstract: Empathy is essential for fostering natural interactions in spoken dialogue systems, as it enables machines to recognize the emotional tone of human spee

researcharxiv-cs-cl
21 Apr 2026
Model Releases

FregeLogic at SemEval 2026 Task 11: A Hybrid Neuro-Symbolic Architecture for Content-Robust Syllogistic Validity Prediction

DGX agent

arXiv:2604.18328v1 Announce Type: new Abstract: We present FregeLogic, a hybrid neuro-symbolic system for SemEval-2026 Task 11 (Subtask 1), which addresses syllogistic validity prediction while reduci

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Freshness-Aware Prioritized Experience Replay for LLM/VLM Reinforcement Learning

DGX agent

arXiv:2604.16918v1 Announce Type: new Abstract: Reinforcement Learning (RL) has achieved impressive success in post-training Large Language Models (LLMs) and Vision-Language Models (VLMs), with on-pol

model-releasesarxiv-cs-cl
21 Apr 2026
Research

From Attribution to Abstention: Training-Free Attention-Based Auditing for Clinical Summarization

DGX agent

arXiv:2601.16397v2 Announce Type: replace Abstract: Deploying multimodal large language models (MLLMs) for clinical summarization demands not only fluent generation but also transparency about where e

researcharxiv-cs-cl
21 Apr 2026
Research

From Domains to Instances: Dual-Granularity Data Synthesis for LLM Unlearning

DGX agent

arXiv:2601.04278v2 Announce Type: replace Abstract: Although machine unlearning is essential for removing private, harmful, or copyrighted content from LLMs, current benchmarks often fail to faithfull

researcharxiv-cs-cl
21 Apr 2026
Research

From Fallback to Frontline: When Can LLMs be Superior Annotators of Human Perspectives?

DGX agent

arXiv:2604.17968v1 Announce Type: cross Abstract: Although large language models (LLMs) are increasingly used as annotators at scale, they are typically treated as a pragmatic fallback rather than a f

researcharxiv-cs-cl
21 Apr 2026
Model Releases

From Heads to Neurons: Causal Attribution and Steering in Multi-Task Vision-Language Models

DGX agent

arXiv:2604.17941v1 Announce Type: cross Abstract: Recent work has increasingly explored neuron-level interpretation in vision-language models (VLMs) to identify neurons critical to final predictions.

model-releasesarxiv-cs-cl
21 Apr 2026
Research

From Implicit to Explicit: Token-Efficient Logical Supervision for Mathematical Reasoning in LLMs

DGX agent

arXiv:2601.03682v2 Announce Type: replace Abstract: Recent studies reveal that large language models (LLMs) exhibit limited logical reasoning abilities in mathematical problem-solving, instead often r

researcharxiv-cs-cl
21 Apr 2026
Applications

From Legal Text to Executable Decision Models: Evaluating Structured Representations for Legal Decision Model Generation

DGX agent

arXiv:2604.17153v1 Announce Type: new Abstract: Transforming legal text into executable decision logic is a longstanding challenge in legal informatics. With the rise of LLMs, this task has gained ren

applicationsarxiv-cs-cl
21 Apr 2026
Applications

From Static Inference to Dynamic Interaction: A Survey of Streaming Large Language Models

DGX agent

arXiv:2603.04592v3 Announce Type: replace Abstract: Standard Large Language Models (LLMs) are predominantly designed for static inference with pre-defined inputs, which limits their applicability in d

applicationsarxiv-cs-cl
21 Apr 2026
Research

From Verbatim to Gist: Distilling Pyramidal Multimodal Memory via Semantic Information Bottleneck for Long-Horizon Video Agents

DGX agent

arXiv:2603.01455v2 Announce Type: replace-cross Abstract: While multimodal large language models have demonstrated impressive short-term reasoning, they struggle with long-horizon video understanding

researcharxiv-cs-cl
21 Apr 2026
Safety

Function Words as Statistical Cues for Language Learning

DGX agent

arXiv:2601.21191v2 Announce Type: replace Abstract: What statistical properties might support learning abstract grammatical knowledge from linear input? We address this question by examining the stati

safetyarxiv-cs-cl
21 Apr 2026
Applications

FUSE: Ensembling Verifiers with Zero Labeled Data

DGX agent

arXiv:2604.18547v1 Announce Type: cross Abstract: Verification of model outputs is rapidly emerging as a key primitive for both training and real-world deployment of large language models (LLMs). In p

applicationsarxiv-cs-cl
21 Apr 2026
← Previous
1…133134135136137…161
Next →