AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlog
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,794 results
Local Ai

Use Case: Invoice processing with local LLM - Which LLM and hardware requirements?

DGX agent

This discussion explores using Ollama to run large language models locally for invoice processing while maintaining control over data. The thread likely addresses selecting appropriate smaller models

local-air-ollama
12 May 2026
Model Releases
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

UserGPT Technical Report

DGX agent

arXiv:2605.08766v1 Announce Type: cross Abstract: Personalized user understanding from large-scale digital traces remains a fundamental challenge. Traditional user profiling methods rely on discrimina

model-releasesarxiv-cs-cl
12 May 2026
Safety

Users as Annotators: LLM Preference Learning from Comparison Mode

DGX agent

arXiv:2510.13830v2 Announce Type: replace-cross Abstract: Pairwise preference data have played an important role in the alignment of large language models (LLMs). Each sample of such data consists of

safetyarxiv-cs-ai
12 May 2026
Model Releases

Vapi nabs $50M to make voice AI more human

DGX agent

Voice artificial intelligence startup Vapi Inc. said today it has raised 50 million in new funding to change the way people talk to computers, experience phone calls and interact with customer support

model-releasessiliconangle
12 May 2026
Agents

When Agents Say One Thing and Do Another: Validating Elicited Beliefs from LLMs

DGX agent

arXiv:2602.06286v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed in high-stakes settings where good decisions require forming beliefs over the probability of

agentsarxiv-cs-ai
12 May 2026
Model Releases

When Reviews Disagree: Fine-Grained Contradiction Analysis in Scientific Peer Reviews

DGX agent

arXiv:2605.10171v1 Announce Type: cross Abstract: Scientific peer reviews frequently contain conflicting expert judgments, and the increasing scale of conference submissions makes it challenging for A

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

When to Re-Commit: Temporal Abstraction Discovery for Long-Horizon Vision-Language Reasoning

DGX agent

arXiv:2605.09860v1 Announce Type: new Abstract: Long-horizon reasoning requires deciding not only what actions to take, but how deeply to commit before the next observation. We formalize this as commi

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

3 weeks since ml-intern launched and we just hit 1M messages exchanged. that's 3.3 agent-years of ML research in 21 days. 2 months worth of …

DGX agent

3 weeks since ml-intern launched and we just hit 1M messages exchanged. that's 3.3 agent-years of ML research in 21 days. 2 months worth of research every day. 17,383 training jobs total. talk about A

model-releasesclem-delangue--x
11 May 2026
Model Releases

A Hierarchical Ensemble Pipeline for Anomaly Detection in ESA Satellite Telemetry

DGX agent

arXiv:2605.06681v1 Announce Type: cross Abstract: A hierarchical ensemble pipeline is introduced to address anomaly detection in multivariate telemetry data provided by European Space Agency (ESA). Th

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

A Reproducible Multi-Architecture Baseline for Token-Level Chinese Metaphor Identification under the MIPVU Framework

DGX agent

arXiv:2605.07170v1 Announce Type: new Abstract: Metaphor is pervasive in everyday language, yet token-level computational identification of metaphor-related words in Chinese under the MIPVU framework

model-releasesarxiv-cs-cl
11 May 2026
Safety

A Systematic Investigation of The RL-Jailbreaker in LLMs

DGX agent

arXiv:2605.07032v1 Announce Type: cross Abstract: The evolution of generative models from next-token predictors to autonomous engines of complex systems necessitates rigorous safety hardening. Adversa

safetyarxiv-cs-ai
11 May 2026
Research

Adaptive Negative Reinforcement for LLM Reasoning:Dynamically Balancing Correction and Diversity in RLVR

DGX agent

arXiv:2605.07137v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a highly effective method for improving the reasoning abilities of Large Language Mod

researcharxiv-cs-ai
11 May 2026
Model Releases

AgentEscapeBench: Evaluating Out-of-Domain Tool-Grounded Reasoning in LLM Agents

DGX agent

arXiv:2605.07926v1 Announce Type: new Abstract: As LLM-based agents increasingly rely on external tools, it is important to evaluate their ability to sustain tool-grounded reasoning beyond familiar wo

model-releasesarxiv-cs-ai
11 May 2026
Safety

Anisotropic Modality Align

DGX agent

arXiv:2605.07825v1 Announce Type: cross Abstract: Training multimodal large language models has long been limited by the scarcity of high-quality paired multimodal data. Recent studies show that the s

safetyarxiv-cs-cv
11 May 2026
Agents

Are LLM Agents Behaviorally Coherent? Latent Profiles for Social Simulation

DGX agent

arXiv:2509.03736v2 Announce Type: replace Abstract: The impressive capabilities of Large Language Models (LLMs) raise the possibility that synthetic agents can serve as substitutes for real participan

agentsarxiv-cs-ai
11 May 2026
Model Releases

Benchmarked Yet Not Measured -- Generative AI Should be Evaluated Against Real-World Utility

DGX agent

arXiv:2605.06856v1 Announce Type: cross Abstract: Generative AI systems achieve impressive performance on standard benchmarks yet fail to deliver real-world utility, a disconnect we identify across 28

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Bi3: A Biplatform, Bicultural, Biperson Dataset for Social Robot Navigation

DGX agent

arXiv:2605.06863v1 Announce Type: new Abstract: We contribute Bi3, a dataset of social robot navigation among groups of people in a constrained lab space. Compared to prior data collection efforts for

model-releasesarxiv-cs-ro
11 May 2026
Safety

Bias and Uncertainty in LLM-as-a-Judge Estimation

DGX agent

arXiv:2605.06939v1 Announce Type: new Abstract: LLM-as-a-Judge evaluation has become a standard tool for assessing base model performance. However, characterizing performance via the naive estimator,

safetyarxiv-cs-lg
11 May 2026
Model Releases

Breaking Spatial Uniformity: Prior-Guided Mamba with Radial Serialization for Lens Flare Removal

DGX agent

arXiv:2605.07650v1 Announce Type: new Abstract: Lens flares, caused by complex optical aberrations, severely degrade image quality especially in nighttime photography. Although recent restoration meth

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

CoCoReviewBench: A Completeness- and Correctness-Oriented Benchmark for AI Reviewers

DGX agent

arXiv:2605.07905v1 Announce Type: cross Abstract: Despite the rapid development of AI reviewers, evaluating such systems remains challenging: metrics favor overlap with human reviews over correctness.

model-releasesarxiv-cs-ai
11 May 2026
Safety

Conditional generation of antibody sequences with classifier-guided germline-absorbing discrete diffusion

DGX agent

arXiv:2605.06720v1 Announce Type: cross Abstract: Antibody therapeutics are among the most successful modern medicines, yet computationally designing antibodies with desirable binding and developabili

safetyarxiv-cs-ai
11 May 2026
Safety

Confidence-Aware Alignment Makes Reasoning LLMs More Reliable

DGX agent

arXiv:2605.07353v1 Announce Type: new Abstract: Large reasoning models often reach correct answers through flawed intermediate steps, creating a gap between final accuracy and reasoning reliability. E

safetyarxiv-cs-ai
11 May 2026
Research

CRAFT: Forgetting-Aware Intervention-Based Adaptation for Continual Learning

DGX agent

arXiv:2605.05732v2 Announce Type: replace-cross Abstract: Large language models (LLMs) can acquire new capabilities through fine-tuning, but continual adaptation often leads to catastrophic forgetting

researcharxiv-cs-ai
11 May 2026
Model Releases

Data Contamination in Neural Hieroglyphic Translation: A Reproducibility Study

DGX agent

arXiv:2605.07453v1 Announce Type: new Abstract: Ancient and endangered languages pose a unique challenge for NLP: their datasets are inherently scarce, difficult to expand, and built from formulaic co

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Deeply Dual Supervised learning for melanoma recognition

DGX agent

arXiv:2508.01994v2 Announce Type: replace Abstract: As the application of deep learning in dermatology continues to grow, the recognition of melanoma has garnered significant attention, demonstrating

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Echo: KV-Cache-Free Associative Recall with Spectral Koopman Operators

DGX agent

arXiv:2605.06997v1 Announce Type: new Abstract: Long chain-of-thought reasoning and agentic tool-calling produce traces spanning tens of thousands of tokens, yet Transformer KV caches grow linearly wi

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Effective and Memory-Efficient Alternatives to ECC for Reliable Large-Scale DNNs

DGX agent

arXiv:2605.07417v1 Announce Type: cross Abstract: Modern Deep Learning (DL) workloads are increasingly deployed in safety-critical domains, such as automotive systems and hyperscale data centers, wher

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

End-to-end PDDL Planning with Hardcoded and Dynamic Agents

DGX agent

arXiv:2512.09629v2 Announce Type: replace Abstract: We present an end-to-end framework for planning supported by verifiers. An orchestrator receives a human specification written in natural language a

model-releasesarxiv-cs-ai
11 May 2026
Applications

Flexible Routing via Uncertainty Decomposition

DGX agent

arXiv:2605.07805v1 Announce Type: new Abstract: A key strategy for balancing performance and cost in modern machine learning systems is to dynamically route queries to either a low-cost model or a mor

applicationsarxiv-cs-lg
11 May 2026
Safety

From 0-Order Selection to 2-Order Judgment: Combinatorial Hardening Exposes Compositional Failures in Frontier LLMs

DGX agent

arXiv:2605.07268v1 Announce Type: new Abstract: Multiple-choice reasoning benchmarks face dual challenges: rapid saturation from advancing models and data contamination that undermines static evaluati

safetyarxiv-cs-cl
11 May 2026
Model Releases

Gated QKAN-FWP: Scalable Quantum-inspired Sequence Learning

DGX agent

arXiv:2605.06734v1 Announce Type: cross Abstract: Fast Weight Programmers (FWPs) encode temporal dependencies through dynamically updated parameters rather than recurrent hidden states. Quantum FWPs (

model-releasesarxiv-cs-ai
11 May 2026
Local Ai

Geometric Kolmogorov--Arnold Network (GeoKAN)

DGX agent

arXiv:2605.06740v1 Announce Type: cross Abstract: We introduce Geometric Kolmogorov--Arnold Networks (GeoKANs), a family of geometry-aware KAN-type models in which approximation is carried out in lear

local-aiarxiv-cs-ai
11 May 2026
Model Releases

Have Graph -- Will Lift? The Case for Higher-Order Benchmarks

DGX agent

arXiv:2605.07397v1 Announce Type: new Abstract: After a somewhat rocky start, geometry and topology have established a foothold in machine learning. Message passing, either on graphs or higher-order c

model-releasesarxiv-cs-lg
11 May 2026
Industry

Here’s what Mira Murati’s AI company is up to

DGX agent

Thinking Machines, the AI company founded by former OpenAI CTO Mira Murati, announced Monday that it's working on something called 'interaction models.' The idea behind interaction models, according t

industrythe-verge-ai
11 May 2026
Research

How Well Do LLMs Perform on the Simplest Long-Chain Reasoning Tasks: An Empirical Study on the Equivalence Class Problem

DGX agent

arXiv:2605.06882v1 Announce Type: new Abstract: Large Language Models (LLMs) have achieved great improvements in recent years. Nevertheless, it still remains unclear how good LLMs are for reasoning ta

researcharxiv-cs-ai
11 May 2026
Model Releases

HumanNet: Scaling Human-centric Video Learning to One Million Hours

DGX agent

arXiv:2605.06747v1 Announce Type: new Abstract: Progress in embodied intelligence increasingly depends on scalable data infrastructure. While vision and language have scaled with internet corpora, lea

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Interpreting Reinforcement Learning Agents with Susceptibilities

DGX agent

arXiv:2605.08007v1 Announce Type: new Abstract: Susceptibilities are a technique for neural network interpretability that studies the response of posterior expectation values of observables to perturb

model-releasesarxiv-cs-lg
11 May 2026
Research

LaTER: Efficient Test-Time Reasoning via Latent Exploration and Explicit Verification

DGX agent

arXiv:2605.07315v1 Announce Type: new Abstract: Chain-of-thought (CoT) reasoning improves large language models (LLMs) on difficult tasks, but it also makes inference expensive because every intermedi

researcharxiv-cs-cl
11 May 2026
Research

Minimizing Modality Gap from the Input Side: Your Speech LLM Can Be a Prosody-Aware Text LLM

DGX agent

arXiv:2605.05927v2 Announce Type: replace Abstract: Speech large language models (SLMs) are typically built from text large language model (TLM) checkpoints, yet they still suffer from a substantial m

researcharxiv-cs-cl
11 May 2026
Model Releases

Modular Lie Algebraic PDE Control of Multibody Flexible Manipulators

DGX agent

arXiv:2605.06709v1 Announce Type: new Abstract: This paper addresses PDE-based control for flexible multibody robotic systems, presenting a subsystem-based framework for serial manipulators with arbit

model-releasesarxiv-cs-ro
11 May 2026
Research

Multimodal Latent Reasoning via Hierarchical Visual Cues Injection

DGX agent

arXiv:2602.05359v2 Announce Type: replace Abstract: The advancement of multimodal large language models (MLLMs) has enabled impressive perception capabilities. However, their reasoning process often r

researcharxiv-cs-cv
11 May 2026
Model Releases

Muon Dynamics as a Spectral Wasserstein Flow

DGX agent

arXiv:2604.04891v2 Announce Type: replace-cross Abstract: Gradient normalization stabilizes deep-learning optimization, and spectral normalizations are especially natural for matrix-shaped parameter b

model-releasesarxiv-cs-ai
11 May 2026
Safety

PACEvolve++: Improving Test-time Learning for Evolutionary Search Agents

DGX agent

arXiv:2605.07039v1 Announce Type: new Abstract: Large language models have become drivers of evolutionary search, but most systems rely on a fixed, prompt-elicited policy to sample next candidates. Th

safetyarxiv-cs-lg
11 May 2026
Model Releases

PAIR-Former: Budgeted Relational Multi-Instance Learning for Functional miRNA Target Prediction

DGX agent

arXiv:2602.00465v3 Announce Type: replace-cross Abstract: Functional miRNA--mRNA targeting is a large-bag prediction problem where each transcript yields a heavy-tailed pool of candidate target sites

model-releasesarxiv-cs-ai
11 May 2026
Safety

PaT: Planning-after-Trial for Efficient Test-Time Code Generation

DGX agent

arXiv:2605.07248v1 Announce Type: new Abstract: Beyond training-time optimization, scaling test-time computation has emerged as a key paradigm to extend the reasoning capabilities of Large Language Mo

safetyarxiv-cs-cl
11 May 2026
Tutorials

PropSplat: Map-Free RF Field Reconstruction via 3D Gaussian Propagation Splatting

DGX agent

arXiv:2605.08035v1 Announce Type: cross Abstract: Building a site-specific propagation model typically requires either ray-tracing over detailed 3D maps or dense measurement campaigns. Both approaches

tutorialsarxiv-cs-lg
11 May 2026
Hardware

RateQuant: Optimal Mixed-Precision KV Cache Quantization via Rate-Distortion Theory

DGX agent

arXiv:2605.06675v1 Announce Type: cross Abstract: Large language models cache all previously computed key-value (KV) pairs during generation, and this KV cache grows linearly with sequence length, mak

hardwarearxiv-cs-cl
11 May 2026
Model Releases

ReasonSTL: Bridging Natural Language and Signal Temporal Logic via Tool-Augmented Process-Rewarded Learning

DGX agent

arXiv:2605.06483v2 Announce Type: replace Abstract: Signal Temporal Logic (STL) is an expressive formal language for specifying spatio-temporal requirements over real-valued, real-time signals. It has

model-releasesarxiv-cs-ai
11 May 2026
← Previous
1…619620621622623…1371
Next →