AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,223
  • Agents7,699
  • Applications5,506
  • Concepts5
  • Hardware1,889
  • Industry6,186
  • Local Ai5,045
  • Model Releases24,499
  • Research20,615
  • Safety13,633
  • Syntheses17
  • Tools1,677
  • Tutorials3,452

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,223
  • Agents7,699
  • Applications5,506
  • Concepts5
  • Hardware1,889
  • Industry6,186
  • Local Ai5,045
  • Model Releases24,499
  • Research20,615
  • Safety13,633
  • Syntheses17
  • Tools1,677
  • Tutorials3,452

Source
HumanDGX agent

Content type
90,223Total entries
1Added by human
90,222Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
53,226 results
Safety

Anisotropic Modality Align

DGX agent

arXiv:2605.07825v1 Announce Type: cross Abstract: Training multimodal large language models has long been limited by the scarcity of high-quality paired multimodal data. Recent studies show that the s

safetyarxiv-cs-cv
11 May 2026
Agents

Are LLM Agents Behaviorally Coherent? Latent Profiles for Social Simulation

DGX agent

arXiv:2509.03736v2 Announce Type: replace Abstract: The impressive capabilities of Large Language Models (LLMs) raise the possibility that synthetic agents can serve as substitutes for real participan

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
agentsarxiv-cs-ai
11 May 2026
Model Releases

Benchmarked Yet Not Measured -- Generative AI Should be Evaluated Against Real-World Utility

DGX agent

arXiv:2605.06856v1 Announce Type: cross Abstract: Generative AI systems achieve impressive performance on standard benchmarks yet fail to deliver real-world utility, a disconnect we identify across 28

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Bi3: A Biplatform, Bicultural, Biperson Dataset for Social Robot Navigation

DGX agent

arXiv:2605.06863v1 Announce Type: new Abstract: We contribute Bi3, a dataset of social robot navigation among groups of people in a constrained lab space. Compared to prior data collection efforts for

model-releasesarxiv-cs-ro
11 May 2026
Safety

Bias and Uncertainty in LLM-as-a-Judge Estimation

DGX agent

arXiv:2605.06939v1 Announce Type: new Abstract: LLM-as-a-Judge evaluation has become a standard tool for assessing base model performance. However, characterizing performance via the naive estimator,

safetyarxiv-cs-lg
11 May 2026
Model Releases

Breaking Spatial Uniformity: Prior-Guided Mamba with Radial Serialization for Lens Flare Removal

DGX agent

arXiv:2605.07650v1 Announce Type: new Abstract: Lens flares, caused by complex optical aberrations, severely degrade image quality especially in nighttime photography. Although recent restoration meth

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

CoCoReviewBench: A Completeness- and Correctness-Oriented Benchmark for AI Reviewers

DGX agent

arXiv:2605.07905v1 Announce Type: cross Abstract: Despite the rapid development of AI reviewers, evaluating such systems remains challenging: metrics favor overlap with human reviews over correctness.

model-releasesarxiv-cs-ai
11 May 2026
Safety

Conditional generation of antibody sequences with classifier-guided germline-absorbing discrete diffusion

DGX agent

arXiv:2605.06720v1 Announce Type: cross Abstract: Antibody therapeutics are among the most successful modern medicines, yet computationally designing antibodies with desirable binding and developabili

safetyarxiv-cs-ai
11 May 2026
Safety

Confidence-Aware Alignment Makes Reasoning LLMs More Reliable

DGX agent

arXiv:2605.07353v1 Announce Type: new Abstract: Large reasoning models often reach correct answers through flawed intermediate steps, creating a gap between final accuracy and reasoning reliability. E

safetyarxiv-cs-ai
11 May 2026
Research

CRAFT: Forgetting-Aware Intervention-Based Adaptation for Continual Learning

DGX agent

arXiv:2605.05732v2 Announce Type: replace-cross Abstract: Large language models (LLMs) can acquire new capabilities through fine-tuning, but continual adaptation often leads to catastrophic forgetting

researcharxiv-cs-ai
11 May 2026
Model Releases

Data Contamination in Neural Hieroglyphic Translation: A Reproducibility Study

DGX agent

arXiv:2605.07453v1 Announce Type: new Abstract: Ancient and endangered languages pose a unique challenge for NLP: their datasets are inherently scarce, difficult to expand, and built from formulaic co

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Deeply Dual Supervised learning for melanoma recognition

DGX agent

arXiv:2508.01994v2 Announce Type: replace Abstract: As the application of deep learning in dermatology continues to grow, the recognition of melanoma has garnered significant attention, demonstrating

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Echo: KV-Cache-Free Associative Recall with Spectral Koopman Operators

DGX agent

arXiv:2605.06997v1 Announce Type: new Abstract: Long chain-of-thought reasoning and agentic tool-calling produce traces spanning tens of thousands of tokens, yet Transformer KV caches grow linearly wi

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Effective and Memory-Efficient Alternatives to ECC for Reliable Large-Scale DNNs

DGX agent

arXiv:2605.07417v1 Announce Type: cross Abstract: Modern Deep Learning (DL) workloads are increasingly deployed in safety-critical domains, such as automotive systems and hyperscale data centers, wher

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

End-to-end PDDL Planning with Hardcoded and Dynamic Agents

DGX agent

arXiv:2512.09629v2 Announce Type: replace Abstract: We present an end-to-end framework for planning supported by verifiers. An orchestrator receives a human specification written in natural language a

model-releasesarxiv-cs-ai
11 May 2026
Applications

Flexible Routing via Uncertainty Decomposition

DGX agent

arXiv:2605.07805v1 Announce Type: new Abstract: A key strategy for balancing performance and cost in modern machine learning systems is to dynamically route queries to either a low-cost model or a mor

applicationsarxiv-cs-lg
11 May 2026
Safety

From 0-Order Selection to 2-Order Judgment: Combinatorial Hardening Exposes Compositional Failures in Frontier LLMs

DGX agent

arXiv:2605.07268v1 Announce Type: new Abstract: Multiple-choice reasoning benchmarks face dual challenges: rapid saturation from advancing models and data contamination that undermines static evaluati

safetyarxiv-cs-cl
11 May 2026
Model Releases

Gated QKAN-FWP: Scalable Quantum-inspired Sequence Learning

DGX agent

arXiv:2605.06734v1 Announce Type: cross Abstract: Fast Weight Programmers (FWPs) encode temporal dependencies through dynamically updated parameters rather than recurrent hidden states. Quantum FWPs (

model-releasesarxiv-cs-ai
11 May 2026
Local Ai

Geometric Kolmogorov--Arnold Network (GeoKAN)

DGX agent

arXiv:2605.06740v1 Announce Type: cross Abstract: We introduce Geometric Kolmogorov--Arnold Networks (GeoKANs), a family of geometry-aware KAN-type models in which approximation is carried out in lear

local-aiarxiv-cs-ai
11 May 2026
Model Releases

Have Graph -- Will Lift? The Case for Higher-Order Benchmarks

DGX agent

arXiv:2605.07397v1 Announce Type: new Abstract: After a somewhat rocky start, geometry and topology have established a foothold in machine learning. Message passing, either on graphs or higher-order c

model-releasesarxiv-cs-lg
11 May 2026
Research

How Well Do LLMs Perform on the Simplest Long-Chain Reasoning Tasks: An Empirical Study on the Equivalence Class Problem

DGX agent

arXiv:2605.06882v1 Announce Type: new Abstract: Large Language Models (LLMs) have achieved great improvements in recent years. Nevertheless, it still remains unclear how good LLMs are for reasoning ta

researcharxiv-cs-ai
11 May 2026
Model Releases

HumanNet: Scaling Human-centric Video Learning to One Million Hours

DGX agent

arXiv:2605.06747v1 Announce Type: new Abstract: Progress in embodied intelligence increasingly depends on scalable data infrastructure. While vision and language have scaled with internet corpora, lea

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Interpreting Reinforcement Learning Agents with Susceptibilities

DGX agent

arXiv:2605.08007v1 Announce Type: new Abstract: Susceptibilities are a technique for neural network interpretability that studies the response of posterior expectation values of observables to perturb

model-releasesarxiv-cs-lg
11 May 2026
Research

LaTER: Efficient Test-Time Reasoning via Latent Exploration and Explicit Verification

DGX agent

arXiv:2605.07315v1 Announce Type: new Abstract: Chain-of-thought (CoT) reasoning improves large language models (LLMs) on difficult tasks, but it also makes inference expensive because every intermedi

researcharxiv-cs-cl
11 May 2026
Research

Minimizing Modality Gap from the Input Side: Your Speech LLM Can Be a Prosody-Aware Text LLM

DGX agent

arXiv:2605.05927v2 Announce Type: replace Abstract: Speech large language models (SLMs) are typically built from text large language model (TLM) checkpoints, yet they still suffer from a substantial m

researcharxiv-cs-cl
11 May 2026
Model Releases

Modular Lie Algebraic PDE Control of Multibody Flexible Manipulators

DGX agent

arXiv:2605.06709v1 Announce Type: new Abstract: This paper addresses PDE-based control for flexible multibody robotic systems, presenting a subsystem-based framework for serial manipulators with arbit

model-releasesarxiv-cs-ro
11 May 2026
Research

Multimodal Latent Reasoning via Hierarchical Visual Cues Injection

DGX agent

arXiv:2602.05359v2 Announce Type: replace Abstract: The advancement of multimodal large language models (MLLMs) has enabled impressive perception capabilities. However, their reasoning process often r

researcharxiv-cs-cv
11 May 2026
Model Releases

Muon Dynamics as a Spectral Wasserstein Flow

DGX agent

arXiv:2604.04891v2 Announce Type: replace-cross Abstract: Gradient normalization stabilizes deep-learning optimization, and spectral normalizations are especially natural for matrix-shaped parameter b

model-releasesarxiv-cs-ai
11 May 2026
Safety

PACEvolve++: Improving Test-time Learning for Evolutionary Search Agents

DGX agent

arXiv:2605.07039v1 Announce Type: new Abstract: Large language models have become drivers of evolutionary search, but most systems rely on a fixed, prompt-elicited policy to sample next candidates. Th

safetyarxiv-cs-lg
11 May 2026
Model Releases

PAIR-Former: Budgeted Relational Multi-Instance Learning for Functional miRNA Target Prediction

DGX agent

arXiv:2602.00465v3 Announce Type: replace-cross Abstract: Functional miRNA--mRNA targeting is a large-bag prediction problem where each transcript yields a heavy-tailed pool of candidate target sites

model-releasesarxiv-cs-ai
11 May 2026
Safety

PaT: Planning-after-Trial for Efficient Test-Time Code Generation

DGX agent

arXiv:2605.07248v1 Announce Type: new Abstract: Beyond training-time optimization, scaling test-time computation has emerged as a key paradigm to extend the reasoning capabilities of Large Language Mo

safetyarxiv-cs-cl
11 May 2026
Tutorials

PropSplat: Map-Free RF Field Reconstruction via 3D Gaussian Propagation Splatting

DGX agent

arXiv:2605.08035v1 Announce Type: cross Abstract: Building a site-specific propagation model typically requires either ray-tracing over detailed 3D maps or dense measurement campaigns. Both approaches

tutorialsarxiv-cs-lg
11 May 2026
Hardware

RateQuant: Optimal Mixed-Precision KV Cache Quantization via Rate-Distortion Theory

DGX agent

arXiv:2605.06675v1 Announce Type: cross Abstract: Large language models cache all previously computed key-value (KV) pairs during generation, and this KV cache grows linearly with sequence length, mak

hardwarearxiv-cs-cl
11 May 2026
Model Releases

ReasonSTL: Bridging Natural Language and Signal Temporal Logic via Tool-Augmented Process-Rewarded Learning

DGX agent

arXiv:2605.06483v2 Announce Type: replace Abstract: Signal Temporal Logic (STL) is an expressive formal language for specifying spatio-temporal requirements over real-valued, real-time signals. It has

model-releasesarxiv-cs-ai
11 May 2026
Local Ai

Reformulating KV Cache Eviction Problem for Long-Context LLM Inference

DGX agent

arXiv:2605.07234v1 Announce Type: cross Abstract: Large language models (LLMs) support long-context inference but suffer from substantial memory and runtime overhead due to Key-Value (KV) Cache growth

local-aiarxiv-cs-ai
11 May 2026
Model Releases

Rethinking Dense Optical Flow without Test-Time Scaling

DGX agent

arXiv:2605.08000v1 Announce Type: new Abstract: Recent progress in dense optical flow has been driven by increasingly complex architectures and multi-step refinement for test-time scaling. While these

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Revisiting Transformer Layer Parameterization Through Causal Energy Minimization

DGX agent

arXiv:2605.07588v1 Announce Type: cross Abstract: Transformer blocks typically combine multi-head attention (MHA) for token mixing with gated MLPs for token-wise feature transformation, yet many choic

model-releasesarxiv-cs-ai
11 May 2026
Safety

Rollback-Free Stable Brick Structures Generation

DGX agent

arXiv:2605.06947v1 Announce Type: new Abstract: While autoregressive models have advanced 3D generation, creating physically stable brick structures remains a challenge due to the strict requirements

safetyarxiv-cs-lg
11 May 2026
Model Releases

Rubric-Grounded RL: Structured Judge Rewards for Generalizable Reasoning

DGX agent

arXiv:2605.08061v1 Announce Type: new Abstract: We argue that decomposing reward into weighted, verifiable criteria and using an LLM judge to score them provides a partial-credit optimization signal:

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

SCENE: Recognizing Social Norms and Sanctioning in Group Chats

DGX agent

arXiv:2605.07823v1 Announce Type: new Abstract: Online group chats are social spaces with implicit behavior patterns that, when broken, are often met with social sanctioning from the group. The abilit

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

SmellBench: Evaluating LLM Agents on Architectural Code Smell Repair

DGX agent

arXiv:2605.07001v1 Announce Type: cross Abstract: Architectural code smells erode software maintainability and are costly to repair manually, yet unlike localized bugs, they require cross-module reaso

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Target-Aware Data Augmentation for SAT Prediction

DGX agent

arXiv:2605.06931v1 Announce Type: new Abstract: Learning-based approaches to NP-hard problems have shown increasing promise, but their progress is fundamentally constrained by the high cost of generat

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

TEA-Bench: A Systematic Benchmarking of Tool-enhanced Emotional Support Dialogue Agent

DGX agent

arXiv:2601.18700v2 Announce Type: replace Abstract: Emotional Support Conversation requires not only affective expression but also grounded instrumental support to provide trustworthy guidance. Howeve

model-releasesarxiv-cs-ai
11 May 2026
Research

Teacher-Feature Drifting: One-Step Diffusion Distillation with Pretrained Diffusion Representations

DGX agent

arXiv:2605.07327v1 Announce Type: new Abstract: Sampling from pretrained diffusion and flow-matching models typically requires many forward passes to generate diverse and high-fidelity images. Existin

researcharxiv-cs-cv
11 May 2026
Model Releases

Traffic Scenario Orchestration from Language via Constraint Satisfaction

DGX agent

arXiv:2605.06966v1 Announce Type: new Abstract: Autonomous vehicles (AVs) require extensive testing in simulation, but test case generation for driving scenarios is laborious. The desired scenarios ar

model-releasesarxiv-cs-ro
11 May 2026
Model Releases

Trajectory as the Teacher: Few-Step Discrete Flow Matching via Energy-Navigated Distillation

DGX agent

arXiv:2605.07924v1 Announce Type: cross Abstract: Discrete flow matching generates text by iteratively transforming noise tokens into coherent language, but may require hundreds of forward passes. Dis

model-releasesarxiv-cs-ai
11 May 2026
Research

Transformer-Based Wildlife Species Classification from Daily Movement Trajectories

DGX agent

arXiv:2605.06726v1 Announce Type: new Abstract: Inferring the identity of wildlife species from daily movement data alone is a challenging task. We train sequence models on large-scale, 7-species GPS

researcharxiv-cs-lg
11 May 2026
Model Releases

UNCOM: Zero-shot Context-Aware Command Understanding for Tabletop Scenarios

DGX agent

arXiv:2410.06355v3 Announce Type: replace-cross Abstract: This paper presents UNCOM, a novel hybrid framework for interpreting natural human commands in tabletop scenarios. The system integrates multi

model-releasesarxiv-cs-ai
11 May 2026
← Previous
1…512513514515516…1109
Next →