AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,646 results
Model Releases

Graph Neural Networks for Misinformation Detection: Performance-Efficiency Trade-offs

DGX agent

arXiv:2604.08131v1 Announce Type: new Abstract: The rapid spread of online misinformation has led to increasingly complex detection models, including large language models and hybrid architectures. Ho

model-releasesarxiv-cs-cl
10 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

GRASS: Gradient-based Adaptive Layer-wise Importance Sampling for Memory-efficient Large Language Model Fine-tuning

DGX agent

arXiv:2604.07808v1 Announce Type: new Abstract: Full-parameter fine-tuning of large language models is constrained by substantial GPU memory requirements. Low-rank adaptation methods mitigate this cha

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

GroupGPT: A Token-efficient and Privacy-preserving Agentic Framework for Multi-User Chat Assistant

DGX agent

arXiv:2603.01059v3 Announce Type: replace Abstract: Recent advances in large language models (LLMs) have enabled increasingly capable chatbots. However, most existing systems focus on single-user sett

model-releasesarxiv-cs-cl
10 Apr 2026
Research

Guaranteeing Knowledge Integration with Joint Decoding for Retrieval-Augmented Generation

DGX agent

arXiv:2604.08046v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) significantly enhances Large Language Models (LLMs) by providing access to external knowledge. However, current res

researcharxiv-cs-cl
10 Apr 2026
Safety

Guardian-as-an-Advisor: Advancing Next-Generation Guardian Models for Trustworthy LLMs

DGX agent

arXiv:2604.07655v1 Announce Type: cross Abstract: Hard-gated safety checkers often over-refuse and misalign with a vendor's model spec; prevailing taxonomies also neglect robustness and honesty, yield

safetyarxiv-cs-cl
10 Apr 2026
Local Ai

Hallucination Detection and Evaluation of Large Language Model

DGX agent

arXiv:2512.22416v2 Announce Type: replace Abstract: Hallucinations in Large Language Models (LLMs) pose a significant challenge, generating misleading or unverifiable content that undermines trust and

local-aiarxiv-cs-cl
10 Apr 2026
Research

HCRE: LLM-based Hierarchical Classification for Cross-Document Relation Extraction with a Prediction-then-Verification Strategy

DGX agent

arXiv:2604.07937v1 Announce Type: new Abstract: Cross-document relation extraction (RE) aims to identify relations between the head and tail entities located in different documents. Existing approache

researcharxiv-cs-cl
10 Apr 2026
Model Releases

HiCI: Hierarchical Construction-Integration for Long-Context Attention

DGX agent

arXiv:2603.20843v2 Announce Type: replace Abstract: Long-context language modeling is commonly framed as a scalability challenge of token-level attention, yet local-to-global information structuring r

model-releasesarxiv-cs-cl
10 Apr 2026
Safety

How Independent are Large Language Models? A Statistical Framework for Auditing Behavioral Entanglement and Reweighting Verifier Ensembles

DGX agent

arXiv:2604.07650v1 Announce Type: cross Abstract: The rapid growth of the large language model (LLM) ecosystem raises a critical question: are seemingly diverse models truly independent? Shared pretra

safetyarxiv-cs-cl
10 Apr 2026
Safety

How Psychological Learning Paradigms Shaped and Constrained Artificial Intelligence

DGX agent

arXiv:2603.18203v3 Announce Type: replace Abstract: Current artificial intelligence systems struggle with systematic compositional reasoning: the capacity to recombine known components in novel config

safetyarxiv-cs-cl
10 Apr 2026
Safety

HumanLLM: Benchmarking and Improving LLM Anthropomorphism via Human Cognitive Patterns

DGX agent

arXiv:2601.10198v3 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities in reasoning and generation, serving as the foundation for advanced persona s

safetyarxiv-cs-cl
10 Apr 2026
Research

Hybrid CNN-Transformer Architecture for Arabic Speech Emotion Recognition

DGX agent

arXiv:2604.07357v1 Announce Type: new Abstract: Recognizing emotions from speech using machine learning has become an active research area due to its importance in building human-centered applications

researcharxiv-cs-cl
10 Apr 2026
Model Releases

HyperMem: Hypergraph Memory for Long-Term Conversations

DGX agent

arXiv:2604.08256v1 Announce Type: new Abstract: Long-term memory is essential for conversational agents to maintain coherence, track persistent tasks, and provide personalized interactions across exte

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

IatroBench: Pre-Registered Evidence of Iatrogenic Harm from AI Safety Measures

DGX agent

arXiv:2604.07709v1 Announce Type: cross Abstract: Ask a frontier model how to taper six milligrams of alprazolam (psychiatrist retired, ten days of pills left, abrupt cessation causes seizures) and it

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Initialisation Determines the Basin: Efficient Codebook Optimisation for Extreme LLM Quantization

DGX agent

arXiv:2604.08118v1 Announce Type: new Abstract: Additive quantization enables extreme LLM compression with O(1) lookup-table dequantization, making it attractive for edge deployment. Yet at 2-bit prec

model-releasesarxiv-cs-cl
10 Apr 2026
Research

Iterative Formalization and Planning in Partially Observable Environments

DGX agent

arXiv:2505.13126v3 Announce Type: replace-cross Abstract: Using LLMs not to predict plans but to formalize an environment into the Planning Domain Definition Language (PDDL) has been shown to improve

researcharxiv-cs-cl
10 Apr 2026
Model Releases

Kathleen: Oscillator-Based Byte-Level Text Classification Without Tokenization or Attention

DGX agent

arXiv:2604.07969v1 Announce Type: new Abstract: We present Kathleen, a text classification architecture that operates directly on raw UTF-8 bytes using frequency-domain processing -- requiring no toke

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

KEO: Knowledge Extraction on OMIn via Knowledge Graphs and RAG for Safety-Critical Aviation Maintenance

DGX agent

arXiv:2510.05524v2 Announce Type: replace Abstract: We present Knowledge Extraction on OMIn (KEO), a domain-specific knowledge extraction and reasoning framework with large language models (LLMs) in s

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

KV Cache Offloading for Context-Intensive Tasks

DGX agent

arXiv:2604.08426v1 Announce Type: cross Abstract: With the growing demand for long-context LLMs across a wide range of applications, the key-value (KV) cache has become a critical bottleneck for both

model-releasesarxiv-cs-cl
10 Apr 2026
Safety

Large Language Model Post-Training: A Unified View of Off-Policy and On-Policy Learning

DGX agent

arXiv:2604.07941v1 Announce Type: new Abstract: Post-training has become central to turning pretrained large language models (LLMs) into aligned and deployable systems. Recent progress spans supervise

safetyarxiv-cs-cl
10 Apr 2026
Tutorials

Learning is Forgetting: LLM Training As Lossy Compression

DGX agent

arXiv:2604.07569v1 Announce Type: cross Abstract: Despite the increasing prevalence of large language models (LLMs), we still have a limited understanding of how their representational spaces are stru

tutorialsarxiv-cs-cl
10 Apr 2026
Safety

Learning to Negotiate: Multi-Agent Deliberation for Collective Value Alignment in LLMs

DGX agent

arXiv:2603.10476v2 Announce Type: replace Abstract: LLM alignment has progressed in single-agent settings through paradigms such as RL with human feedback (RLHF), while recent work explores scalable a

safetyarxiv-cs-cl
10 Apr 2026
Safety

Learning Who Disagrees: Demographic Importance Weighting for Modeling Annotator Distributions with DiADEM

DGX agent

arXiv:2604.08425v1 Announce Type: cross Abstract: When humans label subjective content, they disagree, and that disagreement is not noise. It reflects genuine differences in perspective shaped by anno

safetyarxiv-cs-cl
10 Apr 2026
Research

Lexical Tone is Hard to Quantize: Probing Discrete Speech Units in Mandarin and Yoruba

DGX agent

arXiv:2604.07467v1 Announce Type: new Abstract: Discrete speech units (DSUs) are derived by quantising representations from models trained using self-supervised learning (SSL). They are a popular repr

researcharxiv-cs-cl
10 Apr 2026
Tutorials

Linear Representations of Hierarchical Concepts in Language Models

DGX agent

arXiv:2604.07886v1 Announce Type: new Abstract: We investigate how and to what extent hierarchical relations (e.g., Japan subset Eastern Asia subset Asia) are encoded in the internal representat

tutorialsarxiv-cs-cl
10 Apr 2026
Applications

LLM-Based Data Generation and Clinical Skills Evaluation for Low-Resource French OSCEs

DGX agent

arXiv:2604.08126v1 Announce Type: new Abstract: Objective Structured Clinical Examinations (OSCEs) are the standard method for assessing medical students' clinical and communication skills through str

applicationsarxiv-cs-cl
10 Apr 2026
Research

LLM Prompt Duel Optimizer: Efficient Label-Free Prompt Optimization

DGX agent

arXiv:2510.13907v3 Announce Type: replace Abstract: Large language models (LLMs) are highly sensitive to prompts, but most automatic prompt optimization (APO) methods assume access to ground-truth ref

researcharxiv-cs-cl
10 Apr 2026
Research

Loop, Think, & Generalize: Implicit Reasoning in Recurrent-Depth Transformers

DGX agent

arXiv:2604.07822v1 Announce Type: new Abstract: We study implicit reasoning, i.e. the ability to combine knowledge or rules within a single forward pass. While transformer-based large language models

researcharxiv-cs-cl
10 Apr 2026
Model Releases

MARCH: Evaluating the Intersection of Ambiguity Interpretation and Multi-hop Inference

DGX agent

arXiv:2509.22750v3 Announce Type: replace Abstract: Real-world multi-hop QA is naturally linked with ambiguity, where a single query can trigger multiple reasoning paths that require independent resol

model-releasesarxiv-cs-cl
10 Apr 2026
Safety

MemReader: From Passive to Active Extraction for Long-Term Agent Memory

DGX agent

arXiv:2604.07877v1 Announce Type: new Abstract: Long-term memory is fundamental for personalized and autonomous agents, yet populating it remains a bottleneck. Existing systems treat memory extraction

safetyarxiv-cs-cl
10 Apr 2026
Agents

Mina: A Multilingual LLM-Powered Legal Assistant Agent for Bangladesh for Empowering Access to Justice

DGX agent

arXiv:2511.08605v3 Announce Type: replace Abstract: Bangladesh's low-income population faces major barriers to affordable legal advice due to complex legal language, procedural opacity, and high costs

agentsarxiv-cs-cl
10 Apr 2026
Model Releases

MinerU2.5-Pro: Pushing the Limits of Data-Centric Document Parsing at Scale

DGX agent

arXiv:2604.04771v2 Announce Type: replace-cross Abstract: Current document parsing methods advance primarily through model architecture innovation, while systematic engineering of training data remain

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Mitigating Distribution Sharpening in Math RLVR via Distribution-Aligned Hint Synthesis and Backward Hint Annealing

DGX agent

arXiv:2604.07747v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) can improve low-k reasoning accuracy while narrowing solution coverage on challenging math que

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

ModeX: Evaluator-Free Best-of-N Selection for Open-Ended Generation

DGX agent

arXiv:2601.02535v2 Announce Type: replace Abstract: Selecting a single high-quality output from multiple stochastic generations remains a fundamental challenge for large language models (LLMs), partic

model-releasesarxiv-cs-cl
10 Apr 2026
Agents

More Capable, Less Cooperative? When LLMs Fail At Zero-Cost Collaboration

DGX agent

arXiv:2604.07821v1 Announce Type: cross Abstract: Large language model (LLM) agents increasingly coordinate in multi-agent systems, yet we lack an understanding of where and why cooperation failures m

agentsarxiv-cs-cl
10 Apr 2026
Applications

OpenSpatial: A Principled Data Engine for Empowering Spatial Intelligence

DGX agent

arXiv:2604.07296v2 Announce Type: replace Abstract: Spatial understanding is a fundamental cornerstone of human-level intelligence. Nonetheless, current research predominantly focuses on domain-specif

applicationsarxiv-cs-cl
10 Apr 2026
Safety

OpenVLThinkerV2: A Generalist Multimodal Reasoning Model for Multi-domain Visual Tasks

DGX agent

arXiv:2604.08539v1 Announce Type: cross Abstract: Group Relative Policy Optimization (GRPO) has emerged as the de facto Reinforcement Learning (RL) objective driving recent advancements in Multimodal

safetyarxiv-cs-cl
10 Apr 2026
Research

Optimal Decay Spectra for Linear Recurrences

DGX agent

arXiv:2604.07658v1 Announce Type: cross Abstract: Linear recurrent models offer linear-time sequence processing but often suffer from suboptimal long-range memory. We trace this to the decay spectrum:

researcharxiv-cs-cl
10 Apr 2026
Agents

ORACLE-SWE: Quantifying the Contribution of Oracle Information Signals on SWE Agents

DGX agent

arXiv:2604.07789v1 Announce Type: cross Abstract: Recent advances in language model (LM) agents have significantly improved automated software engineering (SWE). Prior work has proposed various agenti

agentsarxiv-cs-cl
10 Apr 2026
Agents

OrgForge: A Multi-Agent Simulation Framework for Verifiable Synthetic Corporate Corpora

DGX agent

arXiv:2603.14997v2 Announce Type: replace Abstract: Building and evaluating enterprise AI systems requires synthetic organizational corpora that are internally consistent, temporally structured, and c

agentsarxiv-cs-cl
10 Apr 2026
Research

Paragraph Segmentation Revisited: Towards a Standard Task for Structuring Speech

DGX agent

arXiv:2512.24517v2 Announce Type: replace Abstract: Automatic speech transcripts are often delivered as unstructured word streams that impede readability and repurposing. We recast paragraph segmentat

researcharxiv-cs-cl
10 Apr 2026
Model Releases

PASK: Toward Intent-Aware Proactive Agents with Long-Term Memory

DGX agent

arXiv:2604.08000v1 Announce Type: cross Abstract: Proactivity is a core expectation for AGI. Prior work remains largely confined to laboratory settings, leaving a clear gap in real-world proactive age

model-releasesarxiv-cs-cl
10 Apr 2026
Safety

PEER: Unified Process-Outcome Reinforcement Learning for Structured Empathetic Reasoning

DGX agent

arXiv:2508.09521v2 Announce Type: replace Abstract: Emotional support conversations require more than fluent responses. Supporters need to understand the seeker's situation and emotions, adopt an appr

safetyarxiv-cs-cl
10 Apr 2026
Model Releases

PeReGrINE: Evaluating Personalized Review Fidelity with User Item Graph Context

DGX agent

arXiv:2604.07788v1 Announce Type: cross Abstract: We introduce PeReGrINE, a benchmark and evaluation framework for personalized review generation grounded in graph-structured user--item evidence. PeRe

model-releasesarxiv-cs-cl
10 Apr 2026
Applications

PIArena: A Platform for Prompt Injection Evaluation

DGX agent

arXiv:2604.08499v1 Announce Type: cross Abstract: Prompt injection attacks pose serious security risks across a wide range of real-world applications. While receiving increasing attention, the communi

applicationsarxiv-cs-cl
10 Apr 2026
Model Releases

PIKA: Expert-Level Synthetic Datasets for Post-Training Alignment from Scratch

DGX agent

arXiv:2510.06670v2 Announce Type: replace Abstract: High-quality instruction data is critical for LLM alignment, yet existing open-source datasets often lack efficiency, requiring hundreds of thousand

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Prompt reinforcing for long-term planning of large language models

DGX agent

arXiv:2510.05921v2 Announce Type: replace Abstract: Large language models (LLMs) have achieved remarkable success in a wide range of natural language processing tasks and can be adapted through prompt

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Prune-Quantize-Distill: An Ordered Pipeline for Efficient Neural Network Compression

DGX agent

arXiv:2604.04988v1 Announce Type: cross Abstract: Modern deployment often requires trading accuracy for efficiency under tight CPU and memory constraints, yet common compression proxies such as parame

model-releasesarxiv-cs-cl
10 Apr 2026
← Previous
1…156157158159160
Next →