AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Research

From Gentlemen to Frontiermen: Masculine Formations in English-Language Fiction (1771--1930)

DGX agent

arXiv:2607.03323v1 Announce Type: new Abstract: Masculinity in nineteenth-century fiction is not a single ideal but a field of competing scripts. Drawing on 150 British and American canonical novels f

researcharxiv-cs-cl
7 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

GameEngineBench: Evaluating Coding Agents on Real C++ Runtime Environments

DGX agent

arXiv:2607.03525v1 Announce Type: cross Abstract: Game engines provide real-time simulation, rendering, physics, interaction, networking, and asset pipelines, making them valuable not only for games b

model-releasesarxiv-cs-cl
7 Jul 2026
Safety

Generative Pseudo-Labeling for Pre-Ranking with LLMs

DGX agent

arXiv:2602.20995v2 Announce Type: replace-cross Abstract: Pre-ranking is a critical stage in industrial recommendation systems, tasked with efficiently scoring thousands of recalled items for downstre

safetyarxiv-cs-cl
7 Jul 2026
Model Releases

GRAFT: Grafted Reference Audio for Fine-grained Pronunciation in Zero-shot Text-to-Speech

DGX agent

arXiv:2607.02633v1 Announce Type: cross Abstract: We present GRAFT, a per-word pronunciation conditioning mechanism for text-to-speech neural codec language modeling. Existing systems reach high intel

model-releasesarxiv-cs-cl
7 Jul 2026
Research

GRASP: Graph-Reasoning Aided Survey Planning for High-Fidelity Related Work Generation

DGX agent

arXiv:2607.03709v1 Announce Type: new Abstract: Writing a literature review requires a deep understanding of the relationships among cited papers: how they build on, challenge, or offer alternative pe

researcharxiv-cs-cl
7 Jul 2026
Applications

HiSAC: Hierarchical Sparse Activation Compression for Ultra-long Sequence Modeling in Recommenders

DGX agent

arXiv:2602.21009v2 Announce Type: replace-cross Abstract: Modern recommender systems leverage ultra-long user behavior sequences to capture dynamic preferences, but end-to-end modeling is infeasible i

applicationsarxiv-cs-cl
7 Jul 2026
Tutorials

How Much is Left? LLMs Linearly Encode Their Remaining Output Length

DGX agent

arXiv:2607.05316v1 Announce Type: new Abstract: Large language models generate one token at a time, yet their responses show remarkably consistent length structure: step-by-step solutions converge in

tutorialsarxiv-cs-cl
7 Jul 2026
Tutorials

How to Build Digital Humans? From Priors to Photorealistic Avatars

DGX agent

arXiv:2607.04341v1 Announce Type: cross Abstract: This state-of-the-art report provides an overview of controllable 3D human avatar creation. We describe current 3D avatar systems, which typically con

tutorialsarxiv-cs-cl
7 Jul 2026
Safety

How Utilitarian Are OpenAI's Models Really? Replicating and Reinterpreting Pfeffer, Krugel, and Uhl (2025)

DGX agent

arXiv:2603.22730v2 Announce Type: replace Abstract: Pfeffer, Krugel, and Uhl (2025) report that OpenAI's reasoning model o1-mini produces more utilitarian responses to the trolley problem and footbrid

safetyarxiv-cs-cl
7 Jul 2026
Model Releases

Hyper-KGGen: A Skill-Driven Knowledge Extractor for High-Quality Knowledge Hypergraph Generation

DGX agent

arXiv:2602.19543v2 Announce Type: replace Abstract: Knowledge hypergraphs surpass traditional binary knowledge graphs by encapsulating complex n-ary atomic facts, providing a more comprehensive paradi

model-releasesarxiv-cs-cl
7 Jul 2026
Safety

Identifiability Without Gaussianity: Symbolic World Models and Near-Infinite Temporal Consistency

DGX agent

arXiv:2606.12471v2 Announce Type: replace-cross Abstract: Klindt, LeCun, and Balestriero (arXiv:2605.26379) proved that Joint-Embedding Predictive Architectures (JEPAs) achieve linear identifiability,

safetyarxiv-cs-cl
7 Jul 2026
Safety

Improving LLMs via Validator-to-Generator Alignment

DGX agent

arXiv:2607.02668v1 Announce Type: new Abstract: Large language models are inconsistent: varying prompts or including unrelated information can lead to unexpected changes in model outputs. The generato

safetyarxiv-cs-cl
7 Jul 2026
Model Releases

Is Your Benchmark Still Useful? Dynamic Benchmarking for Code Language Models

DGX agent

arXiv:2503.06643v2 Announce Type: replace-cross Abstract: In this paper, we tackle a critical challenge in model evaluation: how to keep code benchmarks useful when models might have already seen them

model-releasesarxiv-cs-cl
7 Jul 2026
Research

Jointly Improving Dialect Identification and ASR in Indian Languages using Multimodal Feature Fusion

DGX agent

arXiv:2607.02862v1 Announce Type: new Abstract: Automatic Speech Recognition (ASR) and Dialect Identification (DID) are crucial for Indian languages, many of which are low-resource and exhibit signifi

researcharxiv-cs-cl
7 Jul 2026
Safety

Knowing When Not to Answer: Lightweight KB-Aligned OOD Detection for Safe RAG

DGX agent

arXiv:2508.02296v3 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) systems are increasingly deployed in high-stakes domains, where safety depends not only on how a system answers

safetyarxiv-cs-cl
7 Jul 2026
Model Releases

Knowing When to Stop: Predicting Execution-Consistency Convergence in Text-to-SQL

DGX agent

arXiv:2607.03991v1 Announce Type: cross Abstract: Repeated LLM calls are the standard way to estimate how trustworthy a Text-to-SQL result is: run the pipeline multiple times, judge each SQL execution

model-releasesarxiv-cs-cl
7 Jul 2026
Research

Knowledge Knows, Verbalization Tells: Disentangling Latent Directions for Mathematical Solvability in LLMs

DGX agent

arXiv:2607.05013v1 Announce Type: new Abstract: Although LLMs have made significant progress in mathematical reasoning, determining whether a mathematical problem is solvable remains a fundamental yet

researcharxiv-cs-cl
7 Jul 2026
Safety

Lacuna Inc. at SemEval-2026 Task 4: Structurally Gated State-Space Models for Disentangling Narrative Similarity

DGX agent

arXiv:2607.03482v1 Announce Type: new Abstract: In this paper, we present the Invariant-Variant Disentangled State-Space Model (IVD-SSM), our submission to SemEval-2026 Task 4 on Narrative Story Simil

safetyarxiv-cs-cl
7 Jul 2026
Research

Language Models as Higher-Order Planning Formalizers

DGX agent

arXiv:2603.23844v2 Announce Type: replace Abstract: Recent work provides overwhelming evidence that LLMs, even those trained to scale their reasoning trace, quickly deteriorate at planning as problems

researcharxiv-cs-cl
7 Jul 2026
Research

Large-scale dataset of automatically classified rhetorical sections in scientific papers

DGX agent

arXiv:2607.03381v1 Announce Type: cross Abstract: Scientific papers follow rhetorical structures that organize content into sections such as Introduction, Methods, Results, and Discussion. Automatical

researcharxiv-cs-cl
7 Jul 2026
Safety

Latent Visual Cache for Video Reasoning

DGX agent

arXiv:2607.02607v1 Announce Type: cross Abstract: Video reasoning requires Large Multimodal Models (LMMs) to remain grounded in dense evidence, yet existing systems largely adopt 'read-once, generate-

safetyarxiv-cs-cl
7 Jul 2026
Research

Learning from Lost Provenance: Multiple Instance Learning for Cancer Registry Tumor Group Classification

DGX agent

arXiv:2607.03481v1 Announce Type: new Abstract: Modernizing cancer registries with deep learning is opening new opportunities to automate labor-intensive tasks such as the coding of pathology reports.

researcharxiv-cs-cl
7 Jul 2026
Model Releases

Learning When to Attend: Conditional Memory Access for Long-Context LLMs

DGX agent

arXiv:2603.17484v2 Announce Type: replace Abstract: Language models struggle to generalize beyond pretraining context lengths, limiting long-horizon reasoning and retrieval. Continued pretraining on l

model-releasesarxiv-cs-cl
7 Jul 2026
Model Releases

Legible-by-Construction: Attention and End-to-End Transformers

DGX agent

arXiv:2607.04319v1 Announce Type: new Abstract: A companion paper showed that a transformer's feed-forward layer can be rebuilt from explicit fuzzy set operations - intersection, set-difference, and a

model-releasesarxiv-cs-cl
7 Jul 2026
Safety

LLM-based Human Simulations Have Not Yet Been Reliable

DGX agent

arXiv:2501.08579v3 Announce Type: replace Abstract: Large Language Models (LLMs) are increasingly employed for simulating human behaviors across diverse domains. However, our position is that current

safetyarxiv-cs-cl
7 Jul 2026
Model Releases

LLMs Encode Harmfulness and Refusal Separately

DGX agent

arXiv:2507.11878v5 Announce Type: replace Abstract: LLMs are trained to refuse harmful instructions, but do they truly understand harmfulness beyond just refusing? Prior work has shown that LLMs' refu

model-releasesarxiv-cs-cl
7 Jul 2026
Local Ai

LP-SFT: Local-Preserving Supervised Fine-Tuning via Multimodal Entropy Structure

DGX agent

arXiv:2607.04733v1 Announce Type: new Abstract: Supervised fine-tuning (SFT) is the standard approach for adapting pretrained language models to downstream domains, yet it often improves target-domain

local-aiarxiv-cs-cl
7 Jul 2026
Model Releases

LuxSQA: Ask Me in Luxembourgish with TTS-Augmented Spoken Question Answering

DGX agent

arXiv:2607.02763v1 Announce Type: new Abstract: Spoken Question Answering (SQA) remains largely focused on high-resource languages and carefully recorded speech, limiting the reach of speech-LLM metho

model-releasesarxiv-cs-cl
7 Jul 2026
Research

Measuring and Mitigating Post-hoc Rationalization in Reverse Chain-of-Thought Generation

DGX agent

arXiv:2602.14469v3 Announce Type: replace Abstract: Reverse Chain-of-Thought Generation (RCG) synthesizes reasoning traces from query-answer pairs, but it risks producing post-hoc rationalizations: wh

researcharxiv-cs-cl
7 Jul 2026
Model Releases

Mechanism-level routing failure in LLMs over Lean-verified algebraic structures

DGX agent

arXiv:2607.04534v1 Announce Type: new Abstract: We present an empirical study of structural routing failure in large language models (LLMs) over a formally verified algebraic corpus. The task requires

model-releasesarxiv-cs-cl
7 Jul 2026
Model Releases

Memory-Efficient FastText: A Comprehensive Approach Using Double-Array Trie Structures and Mark-Compact Memory Management

DGX agent

arXiv:2506.01254v2 Announce Type: replace Abstract: FastText remains a practical choice for industrial word representation because it can synthesize vectors for out-of-vocabulary words from character

model-releasesarxiv-cs-cl
7 Jul 2026
Local Ai

Memory-Orchestrated Semantic System (MOSS): An Auditable Agentic Memory Architecture

DGX agent

arXiv:2607.04391v1 Announce Type: new Abstract: Long-term memory remains a structural weakness of AI agents. The dominant approach, retrieval-augmented generation (RAG), relies on embedding-based simi

local-aiarxiv-cs-cl
7 Jul 2026
Research

Mental Health Disorder Detection Beyond Social Media: A Systematic Review of Available Datasets

DGX agent

arXiv:2607.03540v1 Announce Type: new Abstract: Detecting mental health disorders in a timely manner is an important societal challenge. NLP and machine learning (ML) methods used to assist with detec

researcharxiv-cs-cl
7 Jul 2026
Applications

MIRAGE: Defending Long-Form RAG Against Misinformation Pollution

DGX agent

arXiv:2607.05069v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) improves factuality by grounding LLMs in external evidence, but real-world retrieval is often polluted: semanticall

applicationsarxiv-cs-cl
7 Jul 2026
Model Releases

MORE: A Multilingual Document Parsing Benchmark and Evaluation

DGX agent

arXiv:2607.02956v1 Announce Type: cross Abstract: Multilingual documents encapsulate rich regional cultures, scientific discoveries, and historical records. Parsing this content into structured, machi

model-releasesarxiv-cs-cl
7 Jul 2026
Model Releases

MTEB-PT: A Text Embedding Benchmark for Brazilian Portuguese

DGX agent

arXiv:2607.04581v1 Announce Type: new Abstract: Text embeddings for Portuguese have no dedicated benchmark: evaluation rests on translated corpora such as English MS MARCO or on thin multilingual cove

model-releasesarxiv-cs-cl
7 Jul 2026
Local Ai

Multi-Large Language Model Orchestrated Severity Assessment of Clinical Records (MOSAIC)

DGX agent

arXiv:2607.05032v1 Announce Type: new Abstract: Background: Disease severity is a multidimensional construct difficult to capture with rule-based approaches in Electronic Healthcare Records (EHR). Age

local-aiarxiv-cs-cl
7 Jul 2026
Research

OpenSIR: Open-Ended Self-Improving Reasoner

DGX agent

arXiv:2511.00602v4 Announce Type: replace Abstract: Recent advances in large language model (LLM) reasoning through reinforcement learning rely on annotated datasets for verifiable rewards, which may

researcharxiv-cs-cl
7 Jul 2026
Model Releases

Optimizing Large Language Models for Causality Assessment in Pharmacovigilance: Developing a Performance Metric as Objective for Bayesian Hyperparameter Optimization

DGX agent

arXiv:2607.03704v1 Announce Type: new Abstract: Background: Growing individual case safety report (ICSR) volumes have intensified demand for scalable automated causality assessment. Large Language Mod

model-releasesarxiv-cs-cl
7 Jul 2026
Research

Ossetic-COT: Designing a morphologically annotated corpus and morphological analyzer for Ossetic

DGX agent

arXiv:2607.04895v1 Announce Type: new Abstract: In this work we present the first morphologically annotated corpus for Iron Ossetic that conforms to the Universal Dependencies schema. The corpus inclu

researcharxiv-cs-cl
7 Jul 2026
Research

PAST-TIDE: Prototype-Anchored Statement Tuning with Topic-Invariant Normalization for Stance Detection

DGX agent

arXiv:2607.04690v1 Announce Type: new Abstract: We introduce PAST-TIDE, our stance detection system addressing both subtasks of the StanceNakba Shared Task at NakbaNLP@LREC-COLING 2026. The main idea

researcharxiv-cs-cl
7 Jul 2026
Research

Pathways of Visual Information Flow in Vision-Language Models

DGX agent

arXiv:2607.03358v1 Announce Type: cross Abstract: We study how visual information is routed in vision-language models (VLMs). Using causal patching on controlled synthetic and natural datasets, we fin

researcharxiv-cs-cl
7 Jul 2026
Research

PraMem: Practice-derived Experiential Memory for Long-horizon Behavior Prediction

DGX agent

arXiv:2607.02881v1 Announce Type: new Abstract: Long-horizon behavior prediction aims to infer a user's next action based on a lengthy historical sequence, playing a crucial role in artificial intelli

researcharxiv-cs-cl
7 Jul 2026
Research

Predicting the Emergence of Induction Heads in Language Model Pretraining

DGX agent

arXiv:2511.16893v3 Announce Type: replace Abstract: Specialized attention heads dubbed induction heads (IHs) have been argued to underlie the remarkable in-context learning capabilities of modern lang

researcharxiv-cs-cl
7 Jul 2026
Model Releases

ProACT: Towards Breakdown-Aware Proactive Agent in Multi-User Collaboration

DGX agent

arXiv:2607.03730v1 Announce Type: new Abstract: Conversational agents are increasingly embedded in human collaborative work, yet they remain fundamentally passive and reactive: they respond to explici

model-releasesarxiv-cs-cl
7 Jul 2026
Agents

Progressive Disclosure for LLM-Maintained Wiki Knowledge Bases: a Preregistered Ablation

DGX agent

arXiv:2607.04576v1 Announce Type: new Abstract: LLM agents increasingly answer questions against knowledge bases they help maintain. A common intuition holds that progressive disclosure, a compact cat

agentsarxiv-cs-cl
7 Jul 2026
Research

Progressive Refinement: An Iterative Pseudo-Labeling Approach for Mandarin-English Code-Switching ASR

DGX agent

arXiv:2607.05224v1 Announce Type: new Abstract: Code-switching (CS), alternating languages within the same utterance, poses significant challenges for automatic speech recognition (ASR) due to limited

researcharxiv-cs-cl
7 Jul 2026
Research

ProLaViT: Learning Progressive Latent Visual Thoughts in Structured Latent Space

DGX agent

arXiv:2607.02907v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable progress but still struggle with complex visual reasoning tasks requiring multi-step

researcharxiv-cs-cl
7 Jul 2026
← Previous
1…3132333435…161
Next →