AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Applications

LegalMidm: Use-Case-Driven Legal Domain Specialization for Korean Large Language Model

DGX agent

arXiv:2604.25297v1 Announce Type: new Abstract: In recent years, the rapid proliferation of open-source large language models (LLMs) has spurred efforts to turn general-purpose models into domain spec

applicationsarxiv-cs-cl
29 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Less Is More: Fast and Accurate Reasoning with Cross-Head Unified Sparse Attention

DGX agent

arXiv:2508.07101v2 Announce Type: replace Abstract: Large reasoning models achieve strong performance through test-time scaling, but this incurs substantial computational overhead due to long decoding

researcharxiv-cs-cl
29 Apr 2026
Agents

Leverage Laws: A Per-Task Framework for Human-Agent Collaboration

DGX agent

arXiv:2604.25040v1 Announce Type: cross Abstract: We propose a per-task leverage ratio for human-agent collaboration: human work displaced by an agent, divided by the human time required to specify th

agentsarxiv-cs-cl
29 Apr 2026
Safety

Libra-VLA: Achieving Learning Equilibrium via Asynchronous Coarse-to-Fine Dual-System

DGX agent

arXiv:2604.24921v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models are a promising paradigm for generalist robotic manipulation by grounding high-level semantic instructions into ex

safetyarxiv-cs-cl
29 Apr 2026
Research

Limited Linguistic Diversity in Embodied AI Datasets

DGX agent

arXiv:2601.03136v2 Announce Type: replace Abstract: Language plays a critical role in Vision-Language-Action (VLA) models, yet the linguistic characteristics of the datasets used to train and evaluate

researcharxiv-cs-cl
29 Apr 2026
Model Releases

LLM-ReSum: A Framework for LLM Reflective Summarization through Self-Evaluation

DGX agent

arXiv:2604.25665v1 Announce Type: new Abstract: Reliable evaluation of large language model (LLM)-generated summaries remains an open challenge, particularly across heterogeneous domains and document

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization

DGX agent

arXiv:2604.25130v1 Announce Type: new Abstract: Evaluating long document summaries remains the primary bottleneck in summarization research. Existing metrics correlate weakly with human judgments and

model-releasesarxiv-cs-cl
29 Apr 2026
Local Ai

Luminol-AIDetect: Fast Zero-shot Machine-Generated Text Detection based on Perplexity under Text Shuffling

DGX agent

arXiv:2604.25860v1 Announce Type: new Abstract: Machine-generated text (MGT) detection requires identifying structurally invariant signals across generation models, rather than relying on model-specif

local-aiarxiv-cs-cl
29 Apr 2026
Safety

MAIC-UI: Making Interactive Courseware with Generative UI

DGX agent

arXiv:2604.25806v1 Announce Type: new Abstract: Creating interactive STEM courseware traditionally requires HTML/CSS/JavaScript expertise, leaving barriers for educators. While generative AI can produ

safetyarxiv-cs-cl
29 Apr 2026
Research

Marco-MoE: Open Multilingual Mixture-of-Expert Language Models with Efficient Upcycling

DGX agent

arXiv:2604.25578v1 Announce Type: new Abstract: We present Marco-MoE, a suite of fully open multilingual sparse Mixture-of-Experts (MoE) models. Marco-MoE features a highly sparse design in which only

researcharxiv-cs-cl
29 Apr 2026
Model Releases

MGSM-Pro: A Simple Strategy for Robust Multilingual Mathematical Reasoning Evaluation

DGX agent

arXiv:2601.21225v2 Announce Type: replace Abstract: Large language models have made substantial progress in mathematical reasoning. However, benchmark development for multilingual evaluation has lagge

model-releasesarxiv-cs-cl
29 Apr 2026
Research

MGTEVAL: An Interactive Platform for Systemtic Evaluation of Machine-Generated Text Detectors

DGX agent

arXiv:2604.25152v1 Announce Type: cross Abstract: We present MGTEVAL, an extensible platform for systematic evaluation of Machine-Generated Text (MGT) detectors. Despite rapid progress in MGT detectio

researcharxiv-cs-cl
29 Apr 2026
Agents

MiMo-Embodied: X-Embodied Foundation Model Technical Report

DGX agent

arXiv:2511.16518v2 Announce Type: replace-cross Abstract: We open-source MiMo-Embodied, the first cross-embodied foundation model to successfully integrate and achieve state-of-the-art performance in

agentsarxiv-cs-cl
29 Apr 2026
Model Releases

Mitigating Coordinate Prediction Bias from Positional Encoding Failures

DGX agent

arXiv:2510.22102v2 Announce Type: replace-cross Abstract: While Multimodal Large Language Models (MLLMs) excel at general vision-language tasks, precise coordinate prediction remains a significant cha

model-releasesarxiv-cs-cl
29 Apr 2026
Research

Modeling Human-Like Color Naming Behavior in Context

DGX agent

arXiv:2604.25674v1 Announce Type: new Abstract: Modeling the emergence of human-like lexicons in computational systems has advanced through the use of interacting neural agents, which simulate both le

researcharxiv-cs-cl
29 Apr 2026
Safety

MolReFlect: Towards In-Context Fine-grained Alignments between Molecules and Texts

DGX agent

arXiv:2411.14721v2 Announce Type: replace Abstract: Molecule discovery is a pivotal research field, impacting everything from medicine to materials. Recently, Large Language Models (LLMs) have been wi

safetyarxiv-cs-cl
29 Apr 2026
Research

Named Entity Recognition of Historical Texts via Large Language Model

DGX agent

arXiv:2508.18090v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have demonstrated remarkable versatility across a wide range of natural language processing tasks and domains. On

researcharxiv-cs-cl
29 Apr 2026
Safety

Navigating Global AI Regulation: A Multi-Jurisdictional Retrieval-Augmented Generation System

DGX agent

arXiv:2604.25448v1 Announce Type: new Abstract: Navigating AI regulation across jurisdictions is increasingly difficult for policymakers, legal professionals, and researchers. To address this, we pres

safetyarxiv-cs-cl
29 Apr 2026
Model Releases

Odysseys: Benchmarking Web Agents on Realistic Long Horizon Tasks

DGX agent

arXiv:2604.24964v1 Announce Type: cross Abstract: Existing web agent benchmarks have largely converged on short, single-site tasks that frontier models are approaching saturation on. However, real wor

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

OMHBench: Benchmarking Balanced and Grounded Omni-Modal Multi-Hop Reasoning

DGX agent

arXiv:2508.16198v3 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have increasingly supported omni-modal processing across text, vision, and speech. However, existing evalua

model-releasesarxiv-cs-cl
29 Apr 2026
Safety

One Refiner to Unlock Them All: Inference-Time Reasoning Elicitation via Reinforcement Query Refinement

DGX agent

arXiv:2604.25444v1 Announce Type: new Abstract: Large Language Models (LLMs) often fail to utilize their latent reasoning capabilities due to a distributional mismatch between ambiguous human inquirie

safetyarxiv-cs-cl
29 Apr 2026
Model Releases

Phase-Associative Memory: Sequence Modeling in Complex Hilbert Space

DGX agent

arXiv:2604.05030v2 Announce Type: replace Abstract: Experiments probing natural language processing by both humans and LLMs suggest that the meaning of a semantic expression is indeterminate prior to

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

PolyKV: A Shared Asymmetrically-Compressed KV Cache Pool for Multi-Agent LLM Inference

DGX agent

arXiv:2604.24971v1 Announce Type: cross Abstract: We present PolyKV, a system in which multiple concurrent inference agents share a single, asymmetrically compressed KV cache pool. Rather than allocat

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Praxy Voice: Voice-Prompt Recovery + BUPS for Commercial-Class Indic TTS from a Frozen Non-Indic Base at Zero Commercial-Training-Data Cost

DGX agent

arXiv:2604.25441v1 Announce Type: cross Abstract: Commercial TTS systems produce near-native Indic audio, but the best open-source bases (Chatterbox, Indic Parler-TTS, IndicF5) trail them on measured

model-releasesarxiv-cs-cl
29 Apr 2026
Research

Principled Detection of Hallucinations in Large Language Models via Multiple Testing

DGX agent

arXiv:2508.18473v3 Announce Type: replace Abstract: While Large Language Models (LLMs) have emerged as powerful foundational models to solve a variety of tasks, they have also been shown to be prone t

researcharxiv-cs-cl
29 Apr 2026
Safety

Progressing beyond Art Masterpieces or Touristic Cliches: how to assess your LLMs for cultural alignment?

DGX agent

arXiv:2604.25654v1 Announce Type: new Abstract: Although the cultural (mis)alignment of Large Language Models (LLMs) has attracted increasing attention -- often framed in terms of cultural bias -- unt

safetyarxiv-cs-cl
29 Apr 2026
Model Releases

PSI-Bench: Towards Clinically Grounded and Interpretable Evaluation of Depression Patient Simulators

DGX agent

arXiv:2604.25840v1 Announce Type: new Abstract: Patient simulators are gaining traction in mental health training by providing scalable exposure to complex and sensitive patient interactions. Simulati

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

PSP: An Interpretable Per-Dimension Accent Benchmark for Indic Text-to-Speech

DGX agent

arXiv:2604.25476v1 Announce Type: cross Abstract: Standard text-to-speech (TTS) evaluation measures intelligibility (WER, CER) and overall naturalness (MOS, UTMOS) but does not quantify accent. A synt

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Quantifying and Mitigating Socially Desirable Responding in LLMs: A Desirability-Matched Graded Forced-Choice Psychometric Study

DGX agent

arXiv:2602.17262v2 Announce Type: replace Abstract: Human self-report questionnaires are increasingly used in NLP to benchmark and audit large language models (LLMs), from persona consistency to safet

model-releasesarxiv-cs-cl
29 Apr 2026
Agents

R^3-SQL: Ranking Reward and Resampling for Text-to-SQL

DGX agent

arXiv:2604.25325v1 Announce Type: cross Abstract: Modern Text-to-SQL systems generate multiple candidate SQL queries and rank them to judge a final prediction. However, existing methods face two limit

agentsarxiv-cs-cl
29 Apr 2026
Model Releases

RbtAct: Rebuttal as Supervision for Actionable Review Feedback Generation

DGX agent

arXiv:2603.09723v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used across the scientific workflow, including to draft peer-review reports. However, many AI-generate

model-releasesarxiv-cs-cl
29 Apr 2026
Agents

Recursive Multi-Agent Systems

DGX agent

arXiv:2604.25917v1 Announce Type: cross Abstract: Recursive or looped language models have recently emerged as a new scaling axis by iteratively refining the same model computation over latent states

agentsarxiv-cs-cl
29 Apr 2026
Applications

RELIC: Evaluating Complex Reasoning via the Recognition of Languages In-Context

DGX agent

arXiv:2506.05205v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used to solve complex tasks where they must retrieve and compose many pieces of in-context information

applicationsarxiv-cs-cl
29 Apr 2026
Research

Rethinking Layer Redundancy in Large Language Models: Calibration Objectives and Search for Depth Pruning

DGX agent

arXiv:2604.24938v1 Announce Type: cross Abstract: Depth pruning improves the inference efficiency of large language models by removing Transformer blocks. Prior work has focused on importance criteria

researcharxiv-cs-cl
29 Apr 2026
Model Releases

Scaling Probabilistic Transformer via Efficient Cross-Scale Hyperparameter Transfer

DGX agent

arXiv:2604.25409v1 Announce Type: new Abstract: Probabilistic Transformer (PT), a white-box probabilistic model for contextual word representation, has demonstrated substantial similarity to standard

model-releasesarxiv-cs-cl
29 Apr 2026
Agents

SciDER: Scientific Data-centric End-to-end Researcher

DGX agent

arXiv:2603.01421v2 Announce Type: replace-cross Abstract: Automated scientific discovery with large language models is transforming the research lifecycle from ideation to experimentation, yet existin

agentsarxiv-cs-cl
29 Apr 2026
Model Releases

SnapMLA: Efficient Long-Context MLA Decoding via Hardware-Aware FP8 Quantized Pipelining

DGX agent

arXiv:2602.10718v3 Announce Type: replace-cross Abstract: While FP8 attention has shown substantial promise in innovations like FlashAttention-3, its integration into the decoding phase of the DeepSee

model-releasesarxiv-cs-cl
29 Apr 2026
Safety

Subliminal Steering: Stronger Encoding of Hidden Signals

DGX agent

arXiv:2604.25783v1 Announce Type: new Abstract: Subliminal learning describes a student language model inheriting a behavioral bias by fine-tuning on seemingly innocuous data generated by a biased tea

safetyarxiv-cs-cl
29 Apr 2026
Research

The Dynamics of Delusion: Modeling Bidirectional False Belief Amplification in Human-Chatbot Dialogue

DGX agent

arXiv:2604.25096v1 Announce Type: new Abstract: There is growing concern that AI chatbots might fuel delusional beliefs in users. Some have suggested that humans and chatbots mutually reinforce false

researcharxiv-cs-cl
29 Apr 2026
Safety

The Russian Legislative Corpus

DGX agent

arXiv:2406.04855v3 Announce Type: replace Abstract: We present a comprehensive corpus of Russian primary and secondary legislation adopted between 1991 and 2025, comprising 304,382 texts (194,425,905

safetyarxiv-cs-cl
29 Apr 2026
Model Releases

The Structured Output Benchmark: A Multi-Source Benchmark for Evaluating Structured Output Quality in Large Language Models

DGX agent

arXiv:2604.25359v1 Announce Type: new Abstract: Large Language Models are increasingly being deployed to extract structured data from unstructured and semi-structured sources: parsing invoices, medica

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

The Surprising Universality of LLM Outputs: A Real-Time Verification Primitive

DGX agent

arXiv:2604.25634v1 Announce Type: cross Abstract: We report a striking statistical regularity in frontier LLM outputs that enables a CPU-only scoring primitive running at 2.6 microseconds per token, w

model-releasesarxiv-cs-cl
29 Apr 2026
Safety

Thinking About Thinking: Evaluating Reasoning in Post-Trained Language Models

DGX agent

arXiv:2510.16340v2 Announce Type: replace Abstract: Recent advances in post-training techniques have endowed Large Language Models (LLMs) with enhanced capabilities for tackling complex, logic-intensi

safetyarxiv-cs-cl
29 Apr 2026
Safety

Three Models of RLHF Annotation: Extension, Evidence, and Authority

DGX agent

arXiv:2604.25895v1 Announce Type: cross Abstract: Preference-based alignment methods, most prominently Reinforcement Learning with Human Feedback (RLHF), use the judgments of human annotators to shape

safetyarxiv-cs-cl
29 Apr 2026
Safety

TouchAI: Exploring human-AI perceptual alignment in touch through language model representations

DGX agent

arXiv:2406.06587v2 Announce Type: replace Abstract: Aligning large language models (LLMs) behaviour with human intent is critical for future AI. An important yet often overlooked aspect of this alignm

safetyarxiv-cs-cl
29 Apr 2026
Research

Toward a Functional Geometric Algebra for Natural Language Semantics

DGX agent

arXiv:2604.25902v1 Announce Type: new Abstract: Distributional and neural approaches to natural language semantics have been built almost exclusively on conventional linear algebra: vectors, matrices,

researcharxiv-cs-cl
29 Apr 2026
Research

Toward Multimodal Conversational AI for Age-Related Macular Degeneration

DGX agent

arXiv:2604.25720v1 Announce Type: cross Abstract: Despite strong performance of deep learning models in retinal disease detection, most systems produce static predictions without clinical reasoning or

researcharxiv-cs-cl
29 Apr 2026
Safety

Unrequited Emotions: Investigating the Gaps in Motivation and Practice in Speech Emotion Recognition Research

DGX agent

arXiv:2604.25776v1 Announce Type: new Abstract: Critical analyses of emotion recognition technology have raised ethical concerns around task validity and potential downstream impacts, urging researche

safetyarxiv-cs-cl
29 Apr 2026
← Previous
1…119120121122123…161
Next →