AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Tutorials

Neural Garbage Collection: Learning to Forget while Learning to Reason

DGX agent

arXiv:2604.18002v1 Announce Type: new Abstract: Chain-of-thought reasoning has driven striking advances in language model capability, yet every reasoning step grows the KV cache, creating a bottleneck

tutorialsarxiv-cs-lg
21 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

On Different Notions of Redundancy in Conditional-Independence-Based Discovery of Graphical Models

DGX agent

arXiv:2502.08531v3 Announce Type: replace Abstract: Conditional-independence-based discovery uses statistical tests to identify a graphical model that represents the independence structure of variable

researcharxiv-cs-lg
21 Apr 2026
Research

On the Interpolation Effect of Score Smoothing in Diffusion Models

DGX agent

arXiv:2502.19499v3 Announce Type: replace Abstract: Diffusion models have achieved remarkable progress in various domains with an intriguing ability to produce new data that do not exist in the traini

researcharxiv-cs-lg
21 Apr 2026
Model Releases

Precise Debugging Benchmark: Is Your Model Debugging or Regenerating?

DGX agent

arXiv:2604.17338v1 Announce Type: cross Abstract: Unlike code completion, debugging requires localizing faults and applying targeted edits. We observe that frontier LLMs often regenerate correct but o

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

Reward Score Matching: Unifying Reward-based Fine-tuning for Flow and Diffusion Models

DGX agent

arXiv:2604.17415v1 Announce Type: cross Abstract: Reward-based fine-tuning aims to steer a pretrained diffusion or flow-based generative model toward higher-reward samples while remaining close to the

safetyarxiv-cs-cv
21 Apr 2026
Safety

SafeLM: Unified Privacy-Aware Optimization for Trustworthy Federated Large Language Models

DGX agent

arXiv:2604.16606v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed in high-stakes domains, yet a unified treatment of their overlapping safety challenges remains

safetyarxiv-cs-lg
21 Apr 2026
Safety

Safety, Security, and Cognitive Risks in State-Space Models: A Systematic Threat Analysis with Spectral, Stateful, and Capacity Attacks

DGX agent

arXiv:2604.16424v1 Announce Type: cross Abstract: State-Space Models (SSMs) -- structured SSMs (S4, S4D, DSS, S5), selective SSMs (Mamba, Mamba-2), and hybrid architectures (Jamba) -- are deployed in

safetyarxiv-cs-cl
21 Apr 2026
Research

Sparse Feature Coactivation Reveals Causal Semantic Modules in Large Language Models

DGX agent

arXiv:2506.18141v3 Announce Type: replace Abstract: We identify semantically coherent, context-consistent network components in large language models (LLMs) using coactivation of sparse autoencoder (S

researcharxiv-cs-cl
21 Apr 2026
Research

StageMem: Lifecycle-Managed Memory for Language Models

DGX agent

arXiv:2604.16774v1 Announce Type: new Abstract: Long-horizon language model systems increasingly rely on persistent memory, yet many current designs still treat memory primarily as a static store: wri

researcharxiv-cs-cl
21 Apr 2026
Research

Synthetic Data Generation for Training Diversified Commonsense Reasoning Models

DGX agent

arXiv:2603.18361v2 Announce Type: replace Abstract: Conversational agents are required to respond to their users not only with high quality (i.e. commonsense bearing) responses, but also considering m

researcharxiv-cs-cl
21 Apr 2026
Safety

VIBE: Voice-Induced open-ended Bias Evaluation for Large Audio-Language Models via Real-World Speech

DGX agent

arXiv:2604.17248v1 Announce Type: cross Abstract: Large Audio-Language Models (LALMs) are increasingly integrated into daily applications, yet their generative biases remain underexplored. Existing sp

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

When Earth Foundation Models Meet Diffusion: An Application to Land Surface Temperature Super-Resolution

DGX agent

arXiv:2604.16841v1 Announce Type: new Abstract: Land surface temperature (LST) super-resolution is important for environmental monitoring. However, it remains challenging as coarse thermal observation

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

When Visuals Aren't the Problem: Evaluating Vision-Language Models on Misleading Data Visualizations

DGX agent

arXiv:2603.22368v2 Announce Type: replace Abstract: Visualizations help communicate data insights, but deceptive data representations can distort their interpretation and propagate misinformation. Whi

model-releasesarxiv-cs-cv
21 Apr 2026
Agents

XEmbodied: A Foundation Model with Enhanced Geometric and Physical Cues for Large-Scale Embodied Environments

DGX agent

arXiv:2604.18484v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models drive next-generation autonomous systems, but training them requires scalable, high-quality annotations from complex

agentsarxiv-cs-cv
21 Apr 2026
Safety

A Systematic Study of Training-Free Methods for Trustworthy Large Language Models

DGX agent

arXiv:2604.15789v1 Announce Type: new Abstract: As Large Language Models (LLMs) receive increasing attention and are being deployed across various domains, their potential risks, including generating

safetyarxiv-cs-cl
20 Apr 2026
Local Ai

Adapting in the Dark: Efficient and Stable Test-Time Adaptation for Black-Box Models

DGX agent

arXiv:2604.15609v1 Announce Type: cross Abstract: Test-Time Adaptation (TTA) for black-box models accessible only via APIs remains a largely unexplored challenge. Existing approaches such as post-hoc

local-aiarxiv-cs-cv
20 Apr 2026
Model Releases

Characterising LLM-Generated Competency Questions: a Cross-Domain Empirical Study using Open and Closed Models

DGX agent

arXiv:2604.16258v1 Announce Type: new Abstract: Competency Questions (CQs) are a cornerstone of requirement elicitation in ontology engineering. CQs represent requirements as a set of natural language

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

CoMeT: Collaborative Memory Transformer for Efficient Long Context Modeling

DGX agent

arXiv:2602.01766v2 Announce Type: replace-cross Abstract: The quadratic complexity and indefinitely growing key-value (KV) cache of standard Transformers pose a major barrier to long-context processin

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

DiZiNER: Disagreement-guided Instruction Refinement via Pilot Annotation Simulation for Zero-shot Named Entity Recognition

DGX agent

arXiv:2604.15866v1 Announce Type: cross Abstract: Large language models (LLMs) have advanced information extraction (IE) by enabling zero-shot and few-shot named entity recognition (NER), yet their ge

model-releasesarxiv-cs-ai
20 Apr 2026
Research

EchoVLM: Dynamic Mixture-of-Experts Vision-Language Model for Universal Ultrasound Intelligence

DGX agent

arXiv:2509.14977v2 Announce Type: replace Abstract: Ultrasound imaging has become the preferred imaging modality for early cancer screening due to its advantages of non-ionizing radiation, low cost, a

researcharxiv-cs-cv
20 Apr 2026
Model Releases

KWBench: Measuring Unprompted Problem Recognition in Knowledge Work

DGX agent

arXiv:2604.15760v1 Announce Type: new Abstract: We introduce the first version of KWBench (Knowledge Work Bench), a benchmark for unprompted problem recognition in large language models: can an LLM id

model-releasesarxiv-cs-ai
20 Apr 2026
Safety

Large Language Models for Market Research: A Data-augmentation Approach

DGX agent

arXiv:2412.19363v3 Announce Type: replace Abstract: Large Language Models (LLMs) have transformed artificial intelligence by excelling in complex natural language processing tasks. Their ability to ge

safetyarxiv-cs-ai
20 Apr 2026
Research

Large Reasoning Models Are (Not Yet) Multilingual Latent Reasoners

DGX agent

arXiv:2601.02996v2 Announce Type: replace Abstract: Large reasoning models (LRMs) achieve strong performance on mathematical reasoning tasks, often attributed to their capability to generate explicit

researcharxiv-cs-cl
20 Apr 2026
Applications

Modeling of ASD/TD Children's Behaviors in Interaction with a Virtual Social Robot During a Music Education Program Using Deep Neural Networks

DGX agent

arXiv:2604.15314v1 Announce Type: cross Abstract: This research aimed to develop an intelligent system to evaluate performance and extract behavioral models for children with ASD and neurotypical (TD)

applicationsarxiv-cs-ai
20 Apr 2026
Model Releases

Olmo Hybrid: From Theory to Practice and Back

DGX agent

arXiv:2604.03444v3 Announce Type: replace-cross Abstract: Recent work has demonstrated the potential of non-transformer language models, especially linear recurrent neural networks (RNNs) and hybrid m

model-releasesarxiv-cs-cl
20 Apr 2026
Research

Opportunities and Challenges of Large Language Models for Low-Resource Languages in Humanities Research

DGX agent

arXiv:2412.04497v5 Announce Type: replace-cross Abstract: Low-resource languages serve as invaluable repositories of human history, embodying cultural evolution and intellectual diversity. Despite the

researcharxiv-cs-ai
20 Apr 2026
Model Releases

OXtal: An All-Atom Diffusion Model for Organic Crystal Structure Prediction

DGX agent

arXiv:2512.06987v2 Announce Type: replace Abstract: Accurately predicting experimentally realizable 3D molecular crystal structures from their 2D chemical graphs is a long-standing open challenge in c

model-releasesarxiv-cs-lg
20 Apr 2026
Agents

{pi}_{0.7}: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

DGX agent

arXiv:2604.15483v1 Announce Type: new Abstract: We present a new robotic foundation model, called {pi}_{0.7}, that can enable strong out-of-the-box performance in a wide range of scenarios. {pi}_{0.7}

agentsarxiv-cs-lg
20 Apr 2026
Model Releases

PixDLM: A Dual-Path Multimodal Language Model for UAV Reasoning Segmentation

DGX agent

arXiv:2604.15670v1 Announce Type: new Abstract: Reasoning segmentation has recently expanded from ground-level scenes to remote-sensing imagery, yet UAV data poses distinct challenges, including obliq

model-releasesarxiv-cs-cv
20 Apr 2026
Model Releases

Pruning Unsafe Tickets: A Resource-Efficient Framework for Safer and More Robust LLMs

DGX agent

arXiv:2604.15780v1 Announce Type: cross Abstract: Machine learning models are increasingly deployed in real-world applications, but even aligned models such as Mistral and LLaVA still exhibit unsafe b

model-releasesarxiv-cs-cl
20 Apr 2026
Safety

Reward Weighted Classifier-Free Guidance as Policy Improvement in Autoregressive Models

DGX agent

arXiv:2604.15577v1 Announce Type: cross Abstract: Consider an auto-regressive model that produces outputs x (e.g., answers to questions, molecules) each of which can be summarized by an attribute vect

safetyarxiv-cs-ai
20 Apr 2026
Research

Scalable spatial point process models for forensic footwear analysis

DGX agent

arXiv:2602.07006v2 Announce Type: replace Abstract: Shoe print evidence recovered from crime scenes plays a key role in forensic investigations. By examining shoe prints, investigators can determine d

researcharxiv-cs-cv
20 Apr 2026
Research

SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos

DGX agent

arXiv:2602.05638v3 Announce Type: replace Abstract: While foundation models have advanced surgical video analysis, current approaches rely predominantly on pixel-level reconstruction objectives that w

researcharxiv-cs-cv
20 Apr 2026
Research

Teaching Language Models Mechanistic Explainability Through MechSMILES

DGX agent

arXiv:2512.05722v2 Announce Type: replace Abstract: Chemical reaction mechanisms are the foundation of how chemists evaluate reactivity and feasibility, yet current Computer-Assisted Synthesis Plannin

researcharxiv-cs-lg
20 Apr 2026
Hardware

The threat of analytic flexibility in using large language models to simulate human data

DGX agent

arXiv:2509.13397v3 Announce Type: replace-cross Abstract: Social scientists are now using large language models to create 'silicon samples': synthetic datasets intended to stand in for human responden

hardwarearxiv-cs-ai
20 Apr 2026
Research

VIB-Probe: Detecting and Mitigating Hallucinations in Vision-Language Models via Variational Information Bottleneck

DGX agent

arXiv:2601.05547v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) have demonstrated remarkable progress in multimodal tasks, but remain susceptible to hallucinations, where gener

researcharxiv-cs-ai
20 Apr 2026
Research

AlphaCNOT: Learning CNOT Minimization with Model-Based Planning

DGX agent

arXiv:2604.13812v1 Announce Type: new Abstract: Quantum circuit optimization is a central task in Quantum Computing, as current Noisy Intermediate Scale Quantum devices suffer from error propagation t

researcharxiv-cs-ai
17 Apr 2026
Model Releases

AnimationBench: Are Video Models Good at Character-Centric Animation?

DGX agent

arXiv:2604.15299v1 Announce Type: new Abstract: Video generation has advanced rapidly, with recent methods producing increasingly convincing animated results. However, existing benchmarks-largely desi

model-releasesarxiv-cs-cv
17 Apr 2026
Research

Bayesian-LoRA: Probabilistic Low-Rank Adaptation of Large Language Models

DGX agent

arXiv:2601.21003v2 Announce Type: replace Abstract: Large Language Models usually put more emphasis on accuracy and therefore, will guess even when not certain about the prediction, which is especiall

researcharxiv-cs-ai
17 Apr 2026
Research

Compressing Sequences in the Latent Embedding Space: K-Token Merging for Large Language Models

DGX agent

arXiv:2604.15153v1 Announce Type: new Abstract: Large Language Models (LLMs) incur significant computational and memory costs when processing long prompts, as full self-attention scales quadratically

researcharxiv-cs-cl
17 Apr 2026
Research

Dark & Stormy: Modeling Humor in Sentences from the Bulwer-Lytton Fiction Contest

DGX agent

arXiv:2510.24538v2 Announce Type: replace Abstract: Textual humor is enormously diverse and computational studies need to account for this range, including intentionally bad humor. In this paper, we c

researcharxiv-cs-cl
17 Apr 2026
Safety

Exploration and Exploitation Errors Are Measurable for Language Model Agents

DGX agent

arXiv:2604.13151v1 Announce Type: new Abstract: Language Model (LM) agents are increasingly used in complex open-ended decision-making tasks, from AI coding to physical AI. A core requirement in these

safetyarxiv-cs-ai
17 Apr 2026
Model Releases

ImplicitMemBench: Measuring Unconscious Behavioral Adaptation in Large Language Models

DGX agent

arXiv:2604.08064v2 Announce Type: replace Abstract: Existing memory benchmarks for LLM agents evaluate explicit recall of facts, yet overlook implicit memory where experience becomes automated behavio

model-releasesarxiv-cs-ai
17 Apr 2026
Applications

IUQ: Interrogative Uncertainty Quantification for Long-Form Large Language Model Generation

DGX agent

arXiv:2604.15109v1 Announce Type: new Abstract: Despite the rapid advancement of Large Language Models (LLMs), uncertainty quantification in LLM generation is a persistent challenge. Although recent a

applicationsarxiv-cs-cl
17 Apr 2026
Local Ai

Keep It CALM: Toward Calibration-Free Kilometer-Level SLAM with Visual Geometry Foundation Models via an Assistant Eye

DGX agent

arXiv:2604.14795v1 Announce Type: new Abstract: Visual Geometry Foundation Models (VGFMs) demonstrate remarkable zero-shot capabilities in local reconstruction. However, deploying them for kilometer-l

local-aiarxiv-cs-ro
17 Apr 2026
Safety

Model-Free Assessment of Simulator Fidelity via Quantile Curves

DGX agent

arXiv:2512.05024v3 Announce Type: replace-cross Abstract: As generative AI models are increasingly used to simulate real-world systems, quantifying the ``sim-to-real'' gap is critical. For each input

safetyarxiv-cs-lg
17 Apr 2026
Safety

Modeling LLM Unlearning as an Asymmetric Two-Task Learning Problem

DGX agent

arXiv:2604.14808v1 Announce Type: new Abstract: Machine unlearning for large language models (LLMs) aims to remove targeted knowledge while preserving general capability. In this paper, we recast LLM

safetyarxiv-cs-cl
17 Apr 2026
Safety

The PICCO Framework for Large Language Model Prompting: A Taxonomy and Reference Architecture for Prompt Structure

DGX agent

arXiv:2604.14197v1 Announce Type: new Abstract: Large language model (LLM) performance depends heavily on prompt design, yet prompt construction is often described and applied inconsistently. Our purp

safetyarxiv-cs-cl
17 Apr 2026
← Previous
1…142143144145146…1030
Next →