AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,316
  • Agents7,714
  • Applications5,508
  • Concepts5
  • Hardware1,905
  • Industry6,191
  • Local Ai5,052
  • Model Releases24,539
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,678
  • Tutorials3,456

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,316
  • Agents7,714
  • Applications5,508
  • Concepts5
  • Hardware1,905
  • Industry6,191
  • Local Ai5,052
  • Model Releases24,539
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,678
  • Tutorials3,456

Source
HumanDGX agent

Content type
AllBlog
90,316Total entries
1Added by human
90,315Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,172 results
Applications

QuITE: Query-Based Irregular Time Series Embedding

DGX agent

arXiv:2605.28166v1 Announce Type: cross Abstract: Irregular Multivariate Time Series (IMTS) are common in practice, yet their irregular sampling complicates effective modeling. Existing approaches typ

applicationsarxiv-cs-ai
28 May 2026
Model Releases

Relevant Is Not Warranted: Evidence-Force Calibration for Cited RAG

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2605.28044v1 Announce Type: new Abstract: Cited RAG evaluation often treats visible sources as a grounding signal, but a real, topically relevant citation can still under-warrant the attached wo

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

ReSAE: Residualized Sparse Autoencoders for Multi-Layer Transformer Interventions

DGX agent

arXiv:2605.27819v1 Announce Type: cross Abstract: Sparse autoencoders are usually trained one layer at a time, even though transformer residual stream activations are strongly coupled across depth. Th

model-releasesarxiv-cs-ai
28 May 2026
Safety

Restoring the Sweet Spot: Pass-Rate Weighted Self-Distillation for LLM Reasoning

DGX agent

arXiv:2605.27765v1 Announce Type: cross Abstract: Self-Distillation Policy Optimization (SDPO) provides dense token-level credit assignment for reinforcement learning with large language models by lev

safetyarxiv-cs-ai
28 May 2026
Safety

Reward Bias Substitution: Single-Axis Bias Mitigations Redirect Optimization Pressure

DGX agent

arXiv:2605.27996v1 Announce Type: new Abstract: Single-axis mitigations of reward-model biases (e.g., reducing proxy reliance on length, sycophancy, or style) can rotate optimization pressure onto cor

safetyarxiv-cs-ai
28 May 2026
Local Ai

ROVER: Routing Object-Centric Visual Evidence for Grounded Multi-Image Reasoning

DGX agent

arXiv:2605.27959v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have increasingly localized and interleaved visual evidence for deliberative reasoning. Grounding-based appro

local-aiarxiv-cs-ai
28 May 2026
Model Releases

SeeGroup: Multi-Layer Depth Estimation of Transparent Surfaces via Self-Determined Grouping

DGX agent

arXiv:2605.28735v1 Announce Type: new Abstract: Transparent objects are common in daily life, and it is important to understand their multilayer depth, including the transparent surface and the object

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

Self-Supervised Online Robot-Agnostic Traversability Estimation for Open-World Environments

DGX agent

arXiv:2605.28442v1 Announce Type: cross Abstract: Self-supervised online traversability estimation enables robots to continuously learn from unlabeled open-world experiences and adapt their navigation

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

Snippet-Driven Supply Chain Discovery with LLMs: Scaling Visibility in China

DGX agent

arXiv:2605.27845v1 Announce Type: cross Abstract: Financial and economic research often relies on structured supply-chain disclosures and commercial databases. In China, supplier--customer disclosure

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Sparse POD Mode Selection and Manifold Dimensionality Reduction with Neural Networks

DGX agent

arXiv:2605.27756v1 Announce Type: cross Abstract: High-performance computing enables simulation of high-dimensional physical systems, but downstream analyses such as inverse problems and control remai

model-releasesarxiv-cs-lg
28 May 2026
Research

Sparse Scheduled Diffusion Guidance for Inverse Problems

DGX agent

arXiv:2603.07860v2 Announce Type: replace Abstract: Pretrained diffusion models are effective priors for Bayesian inverse problems, but posterior sampling with these priors is often costly because dat

researcharxiv-cs-lg
28 May 2026
Model Releases

Stochastic Gradient Descent with Momentum is Algorithmically Stable

DGX agent

arXiv:2605.28517v1 Announce Type: cross Abstract: Stochastic gradient descent with momentum (SGDM) is one of the most widely used optimization algorithms in machine learning. While optimization proper

model-releasesarxiv-cs-ai
28 May 2026
Safety

Structure-Guided Visual Perturbation Neutralization for LVLMs

DGX agent

arXiv:2605.27927v1 Announce Type: new Abstract: Image inputs enable Large Vision Language Models (LVLMs) to perceive fine-grained visual information, but also introduce a pixel-level attack surface th

safetyarxiv-cs-cv
28 May 2026
Model Releases

Thermodynamic properties of chemically disordered compounds via AI-driven estimation of partition function with the PULSE method

DGX agent

arXiv:2605.28594v1 Announce Type: cross Abstract: In this article, we present an improved version of the PULSE method (Partition function Unsupervised Learning Sampling and Evaluation) for estimating

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

this is so funny, training opus 4.7 on business skills makes it misaligned and dishonest 😭

DGX agent

this is so funny, training opus 4.7 on business skills makes it misaligned and dishonest 😭 Learnings from testing Claude Opus 4.8: > Much worse than Opus 4.7 and GPT 5.5 on Vending Bench > More aligne

model-releasesemad-mostaque--x
28 May 2026
Model Releases

Using Zero-Shot LLM-Generated Survey Data for Geographically Explicit Population Synthesis

DGX agent

arXiv:2605.27401v1 Announce Type: cross Abstract: There is a growing interest in utilizing synthetic populations for a diverse range of applications. At the same time, we are witnessing a tremendous g

model-releasesarxiv-cs-ai
28 May 2026
Safety

VCap: Hypergeometric Rewards for Weak-to-Strong Visual Captioning

DGX agent

arXiv:2605.28023v1 Announce Type: cross Abstract: Visual captioning requires models to capture visual content faithfully while minimizing both omission and hallucination. As the dominant paradigm for

safetyarxiv-cs-ai
28 May 2026
Model Releases

Verifiable Benchmarking of Long-Horizon Spatial Biology

DGX agent

arXiv:2605.28065v1 Announce Type: new Abstract: AI agents are increasingly useful for biological data analysis, but existing benchmarks mostly test broad biological knowledge, executable workflows, or

model-releasesarxiv-cs-ai
28 May 2026
Research

VidPrism: Heterogeneous Mixture of Experts for Image-to-Video Transfer

DGX agent

arXiv:2605.28229v1 Announce Type: cross Abstract: With the rapid development of pre-training technologies, adapting large-scale Vision-Language Models (VLMs) for video understanding ie image-to-video

researcharxiv-cs-ai
28 May 2026
Model Releases

VITAL: Visual-Semantic Dual Supervision for Enhanced and Interpretable Latent Reasoning in Medical MLLMs

DGX agent

arXiv:2605.28422v1 Announce Type: cross Abstract: Latent reasoning enables reasoning over continuous hidden states rather than explicit tokens, avoiding the language bottleneck and inference overhead

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

A newly released AI tool has generated an atlas of more than one billion predicted protein structures and billions more protein sequences. h…

DGX agent

DeepMind's AlphaFold3 and related tools have generated a comprehensive atlas containing over one billion predicted protein structures and additional billions of protein sequences, representing a major

model-releasesyann-lecun--x
27 May 2026
Model Releases

AgentSociety: Incentivizing Agentic Social Intelligence

DGX agent

arXiv:2605.26203v1 Announce Type: cross Abstract: The success of deployed agents relies on their ability to handle open-ended user requests using their inherent capabilities, not only in solving reque

model-releasesarxiv-cs-ai
27 May 2026
Research

Amortized Factor Inference Networks for Posterior Inference

DGX agent

arXiv:2605.26419v1 Announce Type: new Abstract: Amortized inference promises fast test-time Bayesian inference, but existing methods are inherently tied to fixed models. Extending amortization to unse

researcharxiv-cs-lg
27 May 2026
Model Releases

An End-to-End Learning Approach for Solving Capacitated Location-Routing Problems

DGX agent

arXiv:2511.02525v2 Announce Type: replace-cross Abstract: The capacitated location-routing problems (CLRPs) are classical problems in combinatorial optimization, which require simultaneously making lo

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Anchor: Mitigating Artifact Drift in Agent Benchmark Generation

DGX agent

arXiv:2605.26321v1 Announce Type: new Abstract: AI agents are beginning to complete valuable, long-horizon business operations tasks, but training and evaluation environments for enterprise work still

model-releasesarxiv-cs-ai
27 May 2026
Agents

APEX-Searcher: Refining Credit Assignment with Subgoaling for Agentic Retrieval-Augmented Generation

DGX agent

arXiv:2603.13853v3 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) connects large language models (LLMs) to external knowledge, but single-round retrieval is often insuffic

agentsarxiv-cs-ai
27 May 2026
Applications

Behind the MiMo API Price Reduction: The deepest price cut, up to 99%, is for Input (Cache Hit). The core reason is our inference framework …

DGX agent

Behind the MiMo API Price Reduction: The deepest price cut, up to 99%, is for Input (Cache Hit). The core reason is our inference framework now supports hierarchical KV cache optimization for SWA. Pro

applicationsjeremy-howard--x
27 May 2026
Model Releases

Beyond Transfer Accuracy: Faithful Circuits for Controlled Low-Resource Adaptation

DGX agent

arXiv:2601.08146v3 Announce Type: replace-cross Abstract: Existing circuit discovery methods rely on templated tasks with clean counterfactuals, limiting their use on diverse natural text. We adapt Co

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

BeyondSWE: Can Current Code Agent Survive Beyond Single-Repo Bug Fixing?

DGX agent

arXiv:2603.03194v2 Announce Type: replace Abstract: Current code-agent benchmarks primarily evaluate localized issue resolution within a single target repository, leaving under-tested many software en

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

BhashaSetu: A Data-Centric Approach to Low-Resource Machine Translation

DGX agent

arXiv:2605.27050v1 Announce Type: new Abstract: We present BhashaSetu, a linguistically enriched English--Marathi parallel dataset addressing persistent data limitations in low-resource neural machine

model-releasesarxiv-cs-cl
27 May 2026
Agents

Building self-improving tax agents with Codex

DGX agent

This article describes how OpenAI's Codex model can be used to build autonomous tax agents capable of self-improvement through code generation and execution. The work demonstrates using large language

agentsopenai
27 May 2026
Model Releases

Cast a Wider Net: Coordinated Pass@K Policy Optimization for Code Reasoning

DGX agent

arXiv:2605.27000v1 Announce Type: cross Abstract: Repeated sampling with a verifier is the standard way to allocate test-time compute for code generation, with pass@K as the canonical metric. Yet the

model-releasesarxiv-cs-ai
27 May 2026
Research

Conceptual Steganography

DGX agent

arXiv:2605.26537v1 Announce Type: new Abstract: Language Models (LMs) emit Chains-of-Thought (CoTs) that drive much of their capability. However, the same sequence that carries useful reasoning can al

researcharxiv-cs-cl
27 May 2026
Safety

Counteraction-Aware Multi-Teacher On-Policy Distillation for General Capability Recovery with Domain Preservation

DGX agent

arXiv:2605.27115v1 Announce Type: new Abstract: Domain specialization can improve LLM behavior in vertical domains, but often weakens the general capabilities inherited from the original model. Recent

safetyarxiv-cs-ai
27 May 2026
Safety

Curriculum Learning for Safety Alignment

DGX agent

arXiv:2605.26315v1 Announce Type: cross Abstract: Direct Preference Optimisation (DPO) is widely used for safety alignment in large language models. However, prior work shows it is brittle and exhibit

safetyarxiv-cs-ai
27 May 2026
Model Releases

Datacurve releases the DeepSWE coding benchmark, a 113-task test across 91 open-source repositories and five languages, and says GPT-5.5 is the leader at 70% (Michael Nuñez/VentureBeat)

DGX agent

Michael Nuñez / VentureBeat: Datacurve releases the DeepSWE coding benchmark, a 113-task test across 91 open-source repositories and five languages, and says GPT-5.5 is the leader at 70% — For months,

model-releasestechmeme
27 May 2026
Model Releases

DIANOIA: Diagnostic Decomposition and Joint Optimization for Multi-Agent Reasoning

DGX agent

arXiv:2602.08586v3 Announce Type: replace Abstract: Multi-agent LLM systems consistently outperform single-agent baselines, yet practitioners still cannot predict which design works for a new task or

model-releasesarxiv-cs-ai
27 May 2026
Applications

Dynamic Link Prediction with Temporally Enhanced Signed Graph Neural Networks

DGX agent

arXiv:2605.26290v1 Announce Type: new Abstract: Temporal signed networks (TSNs) model the time evolution of cooperative and adversarial relationships that arise in applications such as social media an

applicationsarxiv-cs-lg
27 May 2026
Tutorials

DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding

DGX agent

arXiv:2605.26680v1 Announce Type: cross Abstract: Recent video multimodal large language models (MLLMs) increasingly couple step-by-step reasoning with on-demand visual evidence retrieval, allowing mo

tutorialsarxiv-cs-ai
27 May 2026
Model Releases

ECSEL: Explainable Classification via Signomial Equation Learning

DGX agent

arXiv:2601.21789v2 Announce Type: replace-cross Abstract: We introduce ECSEL, an explainable classification method that learns formal expressions in the form of signomial equations, motivated by the o

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Efficient Prediction of SO(3)-Equivariant Hamiltonian Matrices via SO(2) Local Frames

DGX agent

arXiv:2506.09398v3 Announce Type: replace Abstract: We consider the task of predicting Hamiltonian matrices to accelerate electronic structure calculations, which plays an important role in physics, c

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

EgoProx: Evaluating MLLMs on Egocentric 3D Proximity Reasoning Across a Cognitive Hierarchy

DGX agent

arXiv:2605.24456v2 Announce Type: replace Abstract: Humans constantly reason about 3D proximity, the relations between their body and surrounding objects, to guide perception and action in daily life.

model-releasesarxiv-cs-cv
27 May 2026
Safety

Elias in the Lighthouse, Again? Diagnosing Low Diversity in LLM Stories

DGX agent

arXiv:2605.26492v1 Announce Type: cross Abstract: LLM-generated stories are a popular use case, but they show very low variability. We sample 20,000 total stories from four current models using five p

safetyarxiv-cs-ai
27 May 2026
Model Releases

Enhancing Autonomous Online Intrusion Detection for IoT with Balanced Learning, Reliable Pseudo-Labels, and Lightweight Architectures

DGX agent

arXiv:2605.26166v1 Announce Type: cross Abstract: The rapid proliferation of Internet of Things (IoT) devices has created an urgent demand for adaptive, resource-efficient Intrusion Detection Systems

model-releasesarxiv-cs-ai
27 May 2026
Safety

Ethical Fairness without Demographics in Human-Centered AI

DGX agent

arXiv:2603.13373v3 Announce Type: replace-cross Abstract: In ubiquitous and mobile health systems, computational models infer human states from wearable, behavioral, and physiological sensing data. In

safetyarxiv-cs-ai
27 May 2026
Research

Evaluating the Relevance of Uncertainty Estimators for LLM Hallucination

DGX agent

arXiv:2605.27016v1 Announce Type: cross Abstract: Large language models (LLMs) are prone to hallucinations, i.e., statements unsupported by the input or training data, hindering reliable deployment. I

researcharxiv-cs-ai
27 May 2026
Applications

Function-Valued Causal Influence in Nonlinear Time Series

DGX agent

arXiv:2605.26408v1 Announce Type: new Abstract: Causal discovery in time series is increasingly performed using nonlinear machine-learning models, yet the resulting causal relationships are almost alw

applicationsarxiv-cs-lg
27 May 2026
Model Releases

GEM: Geometric Entropy Mixing for Optimal LLM Data Curation

DGX agent

arXiv:2605.26121v1 Announce Type: cross Abstract: LLM pre-training efficacy increasingly depends on data composition rather than sheer volume. Yet, optimal mixing is hindered by categorization flaws:

model-releasesarxiv-cs-ai
27 May 2026
← Previous
1…702703704705706…1358
Next →