AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,259
  • Agents7,702
  • Applications5,506
  • Concepts5
  • Hardware1,896
  • Industry6,187
  • Local Ai5,048
  • Model Releases24,516
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,677
  • Tutorials3,454

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,259
  • Agents7,702
  • Applications5,506
  • Concepts5
  • Hardware1,896
  • Industry6,187
  • Local Ai5,048
  • Model Releases24,516
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,677
  • Tutorials3,454

Source
HumanDGX agent

Content type
90,259Total entries
1Added by human
90,258Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
53,226 results
Research

NSVQ: Mitigating Codebook Collapse by Stabilizing Encoder Drift in Vector Quantization

DGX agent

arXiv:2606.11363v1 Announce Type: new Abstract: Vector quantization is central to modern generative modeling pipelines, but large-codebook VQ models often suffer from codebook collapse. We identify en

researcharxiv-cs-cv
11 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

OCSVM-Guided Representation Learning for Unsupervised Anomaly Detection

DGX agent

arXiv:2507.21164v2 Announce Type: replace-cross Abstract: Unsupervised anomaly detection (UAD) aims to detect anomalies without labeled data, a necessity in many machine learning applications where an

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

On the Limits of LLM-as-Judge for Scientific Novelty Assessment

DGX agent

arXiv:2606.12071v1 Announce Type: cross Abstract: LLMs are increasingly used to generate and judge scientific ideas. This makes novelty evaluation a central problem. Full idea evaluation is difficult

model-releasesarxiv-cs-ai
11 Jun 2026
Safety

Open Materials Generation with Inference-Time Reinforcement Learning

DGX agent

arXiv:2602.00424v2 Announce Type: replace Abstract: Continuous-time generative models for crystalline materials enable inverse materials design by learning to predict stable crystal structures, but in

safetyarxiv-cs-lg
11 Jun 2026
Model Releases

PCS-UQ: Uncertainty Quantification via the Predictability-Computability-Stability Framework

DGX agent

arXiv:2505.08784v2 Announce Type: replace-cross Abstract: As machine learning (ML) enters high-stakes domains, trustworthy uncertainty quantification (UQ) is essential for safety. In this paper we int

model-releasesarxiv-cs-lg
11 Jun 2026
Tutorials

Position: Stop Anthropomorphizing Intermediate Tokens as Reasoning/Thinking Traces!

DGX agent

arXiv:2504.09762v4 Announce Type: replace Abstract: Intermediate token generation (ITG), where a model produces output before the solution, has become a standard method to improve the performance of l

tutorialsarxiv-cs-ai
11 Jun 2026
Research

Projected random forests and conformal prediction of circular data

DGX agent

arXiv:2410.24145v3 Announce Type: replace-cross Abstract: We apply conformal prediction techniques to regression problems with circular responses, producing prediction sets with adaptive arc length an

researcharxiv-cs-lg
11 Jun 2026
Model Releases

ReMoT: Reinforcement Learning with Motion Contrast Triplets

DGX agent

arXiv:2603.00461v3 Announce Type: replace Abstract: We present ReMoT, a unified training paradigm to systematically address the fundamental shortcomings of VLMs in spatio-temporal consistency -- a cri

model-releasesarxiv-cs-cv
11 Jun 2026
Research

The Long Tail, Not the Front Page: Cold-Start Prediction of Crowd Highlight Salience

DGX agent

arXiv:2606.11654v1 Announce Type: cross Abstract: A social highlighter's most useful signal -- which passages a crowd of readers marks -- exists only for documents people have already read. Can the ag

researcharxiv-cs-cl
11 Jun 2026
Model Releases

The Structural Attention Tax: How Retrieval Format Hijacks In-Context Learning Independent of Content

DGX agent

arXiv:2606.11198v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) systems inject external knowledge to improve LLM outputs, yet the format of injected content -- distinct from its

model-releasesarxiv-cs-ai
11 Jun 2026
Research

Time-Conditioned and Multi-Time Survival Prediction from 2D PET/CT Projections in Lung Cancer

DGX agent

arXiv:2606.12140v1 Announce Type: new Abstract: Accurate prediction of overall survival (OS) from positron emission tomography/computed tomography (PET/CT) can support personalized treatment and follo

researcharxiv-cs-cv
11 Jun 2026
Model Releases

TouchThinker: Scaling Tactile Commonsense Reasoning to the Open World with Large-scale Data and Action-aware Representation

DGX agent

arXiv:2606.11637v1 Announce Type: new Abstract: Touch is a key modality for embodied agents to understand the physical world. Although recent work has incorporated tactile signals into language system

model-releasesarxiv-cs-ai
11 Jun 2026
Tutorials

When is Your LLM Steerable?

DGX agent

arXiv:2606.11599v1 Announce Type: new Abstract: Activation steering offers a lightweight approach to control language models' behavior at inference time, but whether it succeeds or fails heavily depen

tutorialsarxiv-cs-cl
11 Jun 2026
Safety

Alignment Defends LLMs from Property Inference Attacks

DGX agent

arXiv:2606.10217v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly fine-tuned on domain-specific datasets that may contain sensitive, dataset-level properties. Recent work h

safetyarxiv-cs-lg
10 Jun 2026
Safety

AnimaSpark: A Feed-Forward Method for Animating Arbitrary 3D Objects

DGX agent

arXiv:2606.10988v1 Announce Type: new Abstract: While recent advancements in generative AI have substantially accelerated static 3D model creation workflows, the synthesis of category-agnostic 3D anim

safetyarxiv-cs-cv
10 Jun 2026
Model Releases

au-Rec: A Verifiable Benchmark for Agentic Recommender Systems

DGX agent

arXiv:2606.10156v1 Announce Type: cross Abstract: As recommender systems transition toward agentic, multi-turn conversational interfaces, evaluation paradigms have struggled to keep pace. Current benc

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Benchmarking and Exploring the Capabilities of LLMs for Attack Investigations

DGX agent

arXiv:2606.10281v1 Announce Type: cross Abstract: This paper presents AuditBench, a new benchmark dataset for evaluating the capabilities of LLMs at investigating security-related system audit logs. W

model-releasesarxiv-cs-cl
10 Jun 2026
Applications

Data-aware Static Analysis: Improving Detection of Semantic Faults in Machine Learning Code Using Data Characteristics

DGX agent

arXiv:2606.09957v1 Announce Type: cross Abstract: Semantic faults specific to the use of machine learning models are a common problem for machine learning developers, causing suboptimal predictions, h

applicationsarxiv-cs-lg
10 Jun 2026
Model Releases

Divide and Cooperate: Role-Decomposed Multi-Agent LLM Training with Cross-Agent Learning Signals

DGX agent

arXiv:2606.10684v1 Announce Type: cross Abstract: Modern language agents which perform multi-step reasoning have shown strong performance in knowledge-intensive question answering. However, existing a

model-releasesarxiv-cs-ai
10 Jun 2026
Research

Dynamic Linear Attention

DGX agent

arXiv:2606.10650v1 Announce Type: cross Abstract: The scalability of Large Language Models (LLMs) to long contexts is fundamentally constrained by the quadratic complexity of standard attention, motiv

researcharxiv-cs-ai
10 Jun 2026
Model Releases

FreshRetailNet-LT: A Stockout-Annotated Censored Demand Dataset for Latent Demand Recovery and Forecasting in Fresh Retail

DGX agent

arXiv:2505.16319v3 Announce Type: replace Abstract: Accurate demand estimation is critical for the retail business in guiding the inventory and pricing policies of perishable products. However, it fac

model-releasesarxiv-cs-lg
10 Jun 2026
Applications

From Senses to Decisions: The Information Flow of Auditory and Visual Perception in Multimodal LLMs

DGX agent

arXiv:2606.10147v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) can listen and see, but how do audio and visual signals actually travel through the network to shape an answer?

applicationsarxiv-cs-ai
10 Jun 2026
Safety

Gradient-Guided Reward Optimization for Inference-time Alignment

DGX agent

arXiv:2606.09635v1 Announce Type: cross Abstract: Ensuring the reliability of Large Language Models (LLMs) under distribution drift requires inference-time adaptation. While inference-time alignment m

safetyarxiv-cs-lg
10 Jun 2026
Model Releases

How can we assess human-agent interactions? Case studies in software agent design

DGX agent

arXiv:2510.09801v3 Announce Type: replace Abstract: While benchmarks measure the accuracy of LLM-powered agents, they mostly assume full automation, failing to represent the collaborative nature of re

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

JGRA: Jacobian Geometry Robustness Assessment in NISQ Noise-Aware Quantum Neural Networks

DGX agent

arXiv:2606.09964v1 Announce Type: cross Abstract: The NISQ era places stringent constraints on quantum computation, where noise and decoherence fundamentally limit performance. In classical deep learn

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

Latent Guided Sampling for Combinatorial Optimization

DGX agent

arXiv:2506.03672v2 Announce Type: replace-cross Abstract: Combinatorial Optimization problems are widespread in domains such as logistics, manufacturing, and drug discovery, yet their NP-hard nature m

model-releasesarxiv-cs-lg
10 Jun 2026
Safety

Lightweight Latent Reasoning for Narrative Tasks

DGX agent

arXiv:2512.02240v2 Announce Type: replace Abstract: Large language models (LLMs) tackle complex tasks by generating long chains of thought or 'reasoning traces' that act as latent variables in the gen

safetyarxiv-cs-cl
10 Jun 2026
Model Releases

Non-Parametric Structural Priors for Geometry Theorem Prediction

DGX agent

arXiv:2603.04852v2 Announce Type: replace Abstract: Multi-step theorem prediction is a central challenge in geometry problem solving. Existing neural-symbolic approaches rely heavily on supervised par

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

OncoTraj: a public benchmark for longitudinal resistance prediction in EGFR-mutant non-small-cell lung cancer on osimertinib

DGX agent

arXiv:2606.11144v1 Announce Type: new Abstract: Resistance to first-line osimertinib in EGFR-mutant non-small-cell lung cancer (NSCLC) is the canonical example of predictable clonal evolution under th

model-releasesarxiv-cs-lg
10 Jun 2026
Research

One Token per Multimodal Evidence: Latent Memory for Resource-Constrained QA

DGX agent

arXiv:2606.10572v1 Announce Type: new Abstract: External memory effectively grounds large language models (LLMs) and vision-language models (VLMs)-based question answering (QA) in relevant multimodal

researcharxiv-cs-ai
10 Jun 2026
Model Releases

Optimization-based Online Conformal Prediction for Multi-step Forecasting

DGX agent

arXiv:2508.13362v2 Announce Type: replace Abstract: Conformal prediction (CP) is well-suited for uncertainty quantification in time series forecasting due to its distribution-free coverage guarantees.

model-releasesarxiv-cs-lg
10 Jun 2026
Research

Optimizing 2D Input Representations and Sub-phase Fusion Strategies for Differential Diagnosis of Asthma and COPD Using CNN- and GRU-Based Networks

DGX agent

arXiv:2606.10972v1 Announce Type: cross Abstract: This study aims to explore the performance of the VAR model in comparison with mel-frequency cepstral coefficient (MFCC) matrices and log-mel spectrog

researcharxiv-cs-ai
10 Jun 2026
Safety

PADD: Path-Aligned Decompression Distillation for Non-Router Teacher to Guide MoE Student Learning

DGX agent

arXiv:2606.10369v1 Announce Type: new Abstract: As large language models (LLMs) continue to scale, it becomes increasingly challenging to grow model capacity under fixed computation budgets. We propos

safetyarxiv-cs-cl
10 Jun 2026
Model Releases

SCAIL-2: Unifying Controlled Character Animation with End-to-end In-Context Conditioning

DGX agent

arXiv:2606.10804v1 Announce Type: new Abstract: Controlled character animation requires transferring motion from a driving sequence to a reference character. Prior works heavily rely on intermediate r

model-releasesarxiv-cs-cv
10 Jun 2026
Safety

Structure-Preserving Learning Improves Geometry Generalization in Neural PDEs

DGX agent

arXiv:2602.02788v2 Announce Type: replace-cross Abstract: We aim to develop physics foundation models for science and engineering that provide real-time solutions to Partial Differential Equations (PD

safetyarxiv-cs-ai
10 Jun 2026
Model Releases

The 1st PortraitCraft Challenge: A CVPR 2026 Workshop Competition on Portrait Composition Understanding and Generation

DGX agent

arXiv:2606.10894v1 Announce Type: new Abstract: This paper presents an overview of the inaugural PortraitCraft Challenge, held as one of the official competitions at CVPR 2026. The challenge focuses o

model-releasesarxiv-cs-cv
10 Jun 2026
Research

TruthRL: Incentivizing Truthful LLMs via Reinforcement Learning

DGX agent

arXiv:2509.25760v2 Announce Type: replace-cross Abstract: While large language models (LLMs) have demonstrated strong performance on factoid question answering, they are still prone to hallucination a

researcharxiv-cs-ai
10 Jun 2026
Research

Upper Bounds for Local Learning Coefficients of Three-Layer Neural Networks

DGX agent

arXiv:2603.12785v2 Announce Type: replace Abstract: Three-layer neural networks are known to form singular learning models, and their Bayesian asymptotic behavior is governed by the learning coefficie

researcharxiv-cs-lg
10 Jun 2026
Model Releases

Workflow-GYM: Towards Long-Horizon Evaluation of Computer-use Agentic tasks in Real-World Professional Fields

DGX agent

arXiv:2606.11042v1 Announce Type: new Abstract: Recent years have witnessed the rapid evolution of AI agents toward handling increasingly complex, real-world tasks. However, existing benchmarks rarely

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

A Comparative Study of Student Perspectives on Technical Writing Feedback Quality: Evaluating LLMs, SLMs, and Humans in Computer Science Topics

DGX agent

arXiv:2601.11541v2 Announce Type: replace-cross Abstract: To address the scalability of feedback in computer science while mitigating the privacy and cost limitations of commercial Large Language Mode

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

ACTIVE-o3: Empowering MLLMs with Active Perception via Pure Reinforcement Learning

DGX agent

arXiv:2505.21457v2 Announce Type: replace-cross Abstract: Active vision, also known as active perception, refers to actively selecting where and how to look in order to gather task-relevant informatio

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

CHIMERA-Bench: A Benchmark Dataset for Epitope-Specific Antibody Design

DGX agent

arXiv:2603.13431v3 Announce Type: replace-cross Abstract: Computational antibody design has seen rapid methodological progress, with dozens of deep generative methods proposed in the past three years,

model-releasesarxiv-cs-ai
9 Jun 2026
Research

Contemporary AI lacks the imagination to diverge or negate in science

DGX agent

arXiv:2606.08251v1 Announce Type: cross Abstract: Bold projections that artificial intelligence will accelerate scientific discovery have raced ahead of evidence from working scientists, and the field

researcharxiv-cs-ai
9 Jun 2026
Safety

Cranio-Diff: Diffusion-based Cross-domain Craniofacial Reconstruction with 2D X-ray Skull Guidance and Structural Identity Constraints

DGX agent

arXiv:2606.09699v1 Announce Type: new Abstract: The state-of-the-art generative models, such as CycleGAN, Pix2Pix, and diffusion models have demonstrated remarkable performance in the face generation

safetyarxiv-cs-cv
9 Jun 2026
Tutorials

Discovering and decoding latent mean-field structure with variational autoencoders

DGX agent

arXiv:2606.08694v1 Announce Type: cross Abstract: Generative models are increasingly used to capture correlations in many-body systems, but the representations they learn remain largely opaque to phys

tutorialsarxiv-cs-lg
9 Jun 2026
Model Releases

Echo-DM: Ultrasound Marker Removal via Conditional Latent Diffusion and Region-Aware Fusion

DGX agent

arXiv:2606.09378v1 Announce Type: new Abstract: Clinical ultrasound images often contain artificial markers, such as measurement calipers and text, to assist diagnostic interpretation and comparison.

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Emergence World: A Platform for Evaluating Long-Horizon Multi-Agent Autonomy

DGX agent

arXiv:2606.08367v1 Announce Type: cross Abstract: Most evaluations of LLM agents look like exams: a discrete task, a clean environment, a score in minutes or hours. We argue that this approach is mism

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

Emergent alignment and the projectability of ethical personas

DGX agent

arXiv:2606.09475v1 Announce Type: new Abstract: Work on `emergent misalignment' shows that finetuning LLMs on narrow tasks can induce broadly misaligned behavior. This supports the `persona selection'

safetyarxiv-cs-ai
9 Jun 2026
← Previous
1…492493494495496…1109
Next →