AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

87,814Total entries
1Added by human
87,813Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,138 results
26 Jun 2026

Autoformalization of Agent Instructions into Policy-as-Code

Model ReleasesDGX agent

arXiv:2606.26649v1 Announce Type: new Abstract: Agent safety in high-stakes domains requires formal policy enforcement, but most existing approaches either rely on probabilistic guardrails (fine-tuned

b9810

Local AiDGX agent

b9810 is a release build of llama.cpp, an open-source software library that performs inference on various large language models such as Llama. The build identifier follows the project's versioning sch

b9816

Local AiDGX agent

B9816 is a release build number for llama.cpp, an open-source project that enables efficient large language model inference in C/C++ on consumer hardware. This release likely contains bug fixes, perfo

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Bayesian Optimization for General Reaction Conditions
Model ReleasesDGX agent

arXiv:2502.18966v2 Announce Type: replace Abstract: General chemical reaction conditions that achieve consistently high performance across multiple substrates are important for practical applications

Beyond Aesthetics: Quantifying Information Loss in Turbid Scenes

ApplicationsDGX agent

arXiv:2606.26295v1 Announce Type: new Abstract: Visibility in underwater environments degrades rapidly under turbid conditions, yet the effects on computer-vision models remain unclear. This issue is

Beyond Global Divergences: A Local-Mass Perspective on Bayesian Inference

Model ReleasesDGX agent

arXiv:2606.27090v1 Announce Type: cross Abstract: Global objectives, such as KL divergence and ELBO, are widely used in Bayesian inference for measuring distributional discrepancy. This paper studies

Beyond Surface Forms: A Comprehensive, Mechanism-Oriented Taxonomy of Indirect Linguistic Encoding for LLM-Based Coded Language Detection

Model ReleasesDGX agent

arXiv:2606.27314v1 Announce Type: new Abstract: To avoid moderation and surveillance on social media, some users routinely invent indirect linguistic expressions (ILE) that camouflage sensitive meanin

Beyond the Hard Budget: Sparsity Regularizers for More Interpretable Top-k Sparse Autoencoders

ResearchDGX agent

arXiv:2606.27321v1 Announce Type: cross Abstract: Sparse autoencoders (SAEs) have become a leading tool for interpreting the representations of vision foundation models, decomposing their polysemantic

BOWConnect: Parallel Bayesian Optimization over Windows with Learned Local Cost Maps for Sample-Efficient Kinodynamic Motion Planning

Model ReleasesDGX agent

arXiv:2606.27292v1 Announce Type: new Abstract: This paper presents BOWConnect, a bidirectional parallel kinodynamic motion planner that addresses three fundamental limitations of existing sampling-ba

Byzantine-Robust Aggregation for Securing Decentralized Federated Learning

Local AiDGX agent

arXiv:2409.17754v2 Announce Type: replace-cross Abstract: Federated Learning (FL) emerges as a distributed machine learning approach that addresses privacy concerns by training AI models locally on de

Compositionality and the lexicon in evolutionary semantics

ResearchDGX agent

arXiv:2606.27228v1 Announce Type: new Abstract: Formal semantics has shown that sentence meanings arise by recursively composing lexical meanings, yet much of the literature on semantic universals mod

ConvMemory v3: A Validity Context Layer for Conversational Memory via Target-Conditioned Relation Verification

Model ReleasesDGX agent

arXiv:2606.26753v1 Announce Type: new Abstract: Conversational memory retrieval optimizes relevance, yet a retrieved memory can be relevant and simultaneously outdated: a later turn updates, corrects,

Data-driven Machine Learning Cannot Reach Symbolic-level Logical Reasoning -- The Limit of the Scaling Law

Model ReleasesDGX agent

arXiv:2606.26454v1 Announce Type: new Abstract: Sphere neural networks have achieved symbolic level syllogistic reasoning without training data, raising the question of where the limit of the scaling

Decision-Aligned Evaluation of Uncertainty Quantification

Model ReleasesDGX agent

arXiv:2606.26990v1 Announce Type: cross Abstract: Uncertainty estimates in machine learning are typically evaluated using generic metrics such as the negative log-likelihood and expected calibration e

DeCoFlow: Structural Decomposition of Normalizing Flows for Continual Anomaly Detection

Model ReleasesDGX agent

arXiv:2606.26687v1 Announce Type: new Abstract: In industrial environments, new product categories arrive sequentially, requiring continual anomaly detection without access to past data. Normalizing F

DMuon: Efficient Distributed Muon Training with Near-Adam Overhead

ResearchDGX agent

arXiv:2606.27153v1 Announce Type: cross Abstract: Matrix-orthogonalization-based optimizers, exemplified by Muon, have demonstrated strong convergence behavior across a wide range of modern deep learn

Dynamic workflows (generating harnesses on the fly) are a new form of test-time compute. But LLMs aren't great at building them. I often hav…

Model ReleasesDGX agent

Dynamic workflows (generating harnesses on the fly) are a new form of test-time compute. But LLMs aren't great at building them. I often have to steer agents to generate complex patterns. Curious how

EMA-FS: Accelerating GBDT Training via Gain-Informed Feature Screening

Model ReleasesDGX agent

arXiv:2606.26337v1 Announce Type: new Abstract: Gradient Boosted Decision Trees (GBDT), exemplified by LightGBM, spend a dominant fraction of training time -- typically 65-70% -- constructing per-feat

Fast algorithms for learning a Gaussian under halfspace truncation with optimal sample complexity

Model ReleasesDGX agent

arXiv:2606.27298v1 Announce Type: cross Abstract: We study the fundamental problem of learning a high-dimensional Gaussian truncated to an unknown halfspace. Lee, Mehrotra and Zampetakis (FOCS'24) rec

FC-Vision: Real-Time Visibility-Aware Replanning for Occlusion-Free Aerial Target Structure Scanning in Unknown Environments

Model ReleasesDGX agent

arXiv:2602.13720v2 Announce Type: replace Abstract: Autonomous aerial scanning of target structures is crucial for practical applications, requiring online adaptation to unknown obstacles during fligh

Federated Hash Projected Latent Factor Learning

ApplicationsDGX agent

arXiv:2606.26192v1 Announce Type: new Abstract: Hash Learning (HL) is an efficient representation learning approach that maps real-valued data into compact binary representations. Traditional HL metho

FlameVQA: A Physically-Grounded UAV Wildfire VQA Benchmark with Radiometric Thermal Supervision

Model ReleasesDGX agent

arXiv:2606.27128v1 Announce Type: new Abstract: Wildfire monitoring from UAVs requires reliable reasoning over complex aerial scenes, where smoke, scale variation, and occlusions often limit RGB-only

Focusing on What Matters: Saliency-Harnessing Accurate Routing for Diffusion MoE

ResearchDGX agent

arXiv:2606.26938v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) architectures have emerged as a powerful paradigm for scaling diffusion models in visual generation. Recent advancements have f

Forget, Anticipate and Adapt: Test Time Training for Long Videos

TutorialsDGX agent

arXiv:2606.26515v1 Announce Type: new Abstract: Test Time Training (TTT) is a mechanism in which a model adapts to an incoming test-sample by performing some self-supervised (SSL) task and updating it

Fun fact: Dario once delayed the release of GPT-2 back at OpenAI, claiming it was too dangerous

Model ReleasesDGX agent

Fun fact: Dario once delayed the release of GPT-2 back at OpenAI, claiming it was too dangerous Dario fearmogged so hard that global AI progress got halted They could have quietly released it as Opus

Generative AI and Copyright Infringement: A Legal-Technical Analysis of AI Music Generation Systems Under 17 U.S.C. Title 17

Model ReleasesDGX agent

arXiv:2606.26111v1 Announce Type: cross Abstract: Generative artificial intelligence (GenAI) has enabled users to synthesize music with text prompts, combining copyrighted lyrics, AI-composed melodies

GEOALIGN: Geometric Rollout Curation for Robust LLM Reinforcement Learning

SafetyDGX agent

arXiv:2606.26917v1 Announce Type: cross Abstract: Online reinforcement learning is widely used to align large language models (LLMs) with reward signals, yet training can be unstable under noisy or mi

GPT‑5.6 Sol launches with our most robust safety stack yet. We strengthened real-time protections against high-risk cyber activity and repea…

Model ReleasesDGX agent

GPT‑5.6 Sol launches with our most robust safety stack yet. We strengthened real-time protections against high-risk cyber activity and repeated misuse, then spent weeks hardening the system with human

Here is Google Gemini talking about Roger Ebert's review of the 2016 film The Jungle Book. Ebert died in 2013. @GaryMarcus

Model ReleasesDGX agent

This post highlights an apparent error where Google's Gemini AI attributed a film review to Roger Ebert for a 2016 movie, despite Ebert's death in 2013, making such a review impossible. The post, shar

HermesBench full leaderboard coming soon. Stay tuned!

ResearchDGX agent

Nous Research announced an upcoming HermesBench full leaderboard, indicating they are developing or expanding a benchmarking system, likely for evaluating their Hermes model family or related language

hisao{}: A GPU-Native Parallel Optimizer for Multimodal Black-Box Functions via Convergence-Anticonvergence Oscillation

Model ReleasesDGX agent

arXiv:2606.26164v1 Announce Type: new Abstract: Finding all modes of a multimodal black-box function is a fundamental challenge in optimization, Bayesian inference, and scientific computing. Existing

I feel like on X all you hear about is elaborate plans by firms to build their own AI stacks but in my experience companies are full of peop…

Model ReleasesDGX agent

I feel like on X all you hear about is elaborate plans by firms to build their own AI stacks but in my experience companies are full of people who want access to Claude or ChatGPT and are pressuring t

IDEA: Insensitive to Dynamics Mismatch via Effect Alignment for Sim-to-Real Transfer in Multi-Agent Control

SafetyDGX agent

arXiv:2606.26575v1 Announce Type: cross Abstract: Complex multi-agent control tasks remain challenging for traditional rule-based and model-based approaches, motivating the adoption of learning-based

If your benchmark relies on a static dataset or sampling from a static distribution densely known at training time, then it is fundamentally…

Model ReleasesDGX agent

If your benchmark relies on a static dataset or sampling from a static distribution densely known at training time, then it is fundamentally measuring memorization/retrieval. Which might be fine if yo

Instruction Bleed: Cross-Module Interference in Prompt-Composed Agentic Systems

Model ReleasesDGX agent

arXiv:2606.26356v1 Announce Type: new Abstract: Practitioners of prompt-composed agentic systems report a recurring failure mode: editing one prompt module silently shifts the behavior of others despi

Investigating LLM's Problem Solving Capability -- a Study on Statics Questions

ApplicationsDGX agent

arXiv:2606.26103v1 Announce Type: cross Abstract: Large Language Models (LLMs) have rapidly influenced many aspects of society, particularly education, due to their demonstrated ability to complete as

June Launches + Live Q&A Hear from Comfy's CEO @yoland_yan and product leaders Deep Mehta, Alexis Rolland, @jojodecayz , and Matt Miller who…

Model ReleasesDGX agent

June Launches + Live Q&A Hear from Comfy's CEO @yoland_yan and product leaders Deep Mehta, Alexis Rolland, @jojodecayz , and Matt Miller who will walk you through what's new and answer questions live.

KG-TRACE: A Neuro-Symbolic Framework for Mechanistic Grounding in Antimicrobial Resistance Prediction

SafetyDGX agent

arXiv:2606.26179v1 Announce Type: cross Abstract: While WGS-based AMR prediction has reached high accuracy, existing models lack a mechanism to ground neural attributions in established biological pat

Language-Based Digital Twins for Elderly Cognitive Assistance

ApplicationsDGX agent

arXiv:2606.27334v1 Announce Type: new Abstract: Digital twins have emerged as a promising paradigm for personalized healthcare, enabling modeling of individual behavior and health trajectories. In cog

Layered Outer-Loop Control for Disturbance-Robust Multi-Waypoint UAV Arrival

Model ReleasesDGX agent

arXiv:2606.26315v1 Announce Type: new Abstract: Disturbance-robust UAV position control is easy to demonstrate in benign simulations but much harder to make fast in approach, well behaved near the tar

Learning from the Self-future: On-policy Self-distillation for dLLMs

SafetyDGX agent

arXiv:2606.18195v2 Announce Type: replace Abstract: On-policy self-distillation (OPSD) has proven effective for post-training large language models (LLMs), yet its application to diffusion LLMs (dLLMs

Linguistics and Human Brain: A Perspective of Computational Neuroscience

SafetyDGX agent

arXiv:2602.08275v3 Announce Type: replace-cross Abstract: Elucidating the language-brain relationship requires bridging the methodological gap between the abstract theoretical frameworks of linguistic

> mythos is so good at cyber it can't be released also > mythos can't detect 20k fraudulent chinese accounts attacking it

Model ReleasesDGX agent

This post discusses apparent contradictions in claims about Mythos' cybersecurity capabilities, suggesting tension between assertions that it excels at cyber defense versus reports that it failed to d

NaviCache: Test-Time Self-Calibration Caching for Video Generation

SafetyDGX agent

arXiv:2606.26795v1 Announce Type: cross Abstract: Video Diffusion Models (VDMs) is constrained by immense computational costs. While offline calibration-based acceleration suffers from calibration dat

New paper on giving LLM agents experience that improves the weights and stays readable at the same time. Agent-experience methods split into…

Model ReleasesDGX agent

New paper on giving LLM agents experience that improves the weights and stays readable at the same time. Agent-experience methods split into two camps. Externalized natural-language rules stay interpr

not quite all-you-can-eat tokens, but we are working on it

IndustryDGX agent

Sam Altman suggests OpenAI is working toward offering more generous or unlimited token allowances for users, moving beyond current rate-limited pricing models. The post indicates this is an ongoing de

OpenRCA 2.0: From Outcome Labels to Causal Process Supervision

Model ReleasesDGX agent

arXiv:2606.27154v1 Announce Type: new Abstract: Root cause analysis (RCA) poses a holistic test of LLM agentic capabilities, such as long-context understanding, multi-step reasoning, and tool use. How

PathFLIP: Fine-grained Language-Image Pretraining for Versatile Computational Pathology

SafetyDGX agent

arXiv:2512.17621v2 Announce Type: replace Abstract: While Vision-Language Models (VLMs) have achieved notable progress in computational pathology (CPath), the gigapixel scale and spatial heterogeneity

Pianist Transformer: Towards Expressive Piano Performance Rendering via Scalable Self-Supervised Pre-Training

ApplicationsDGX agent

arXiv:2512.02652v2 Announce Type: replace-cross Abstract: Existing methods for expressive music performance rendering, a conditional generation task that aims to generate a human-like performance from

PMDformer: Patch-Mean Decoupling Information Transformer for Long-term Forecasting

ApplicationsDGX agent

arXiv:2606.26549v1 Announce Type: new Abstract: Long-term time series forecasting (LTSF) plays a crucial role in fields such as energy management, finance, and traffic prediction. Transformer-based mo

ProvenAI: Provenance-Native Traces of Evidence in Generated Answers

Model ReleasesDGX agent

arXiv:2606.26449v1 Announce Type: cross Abstract: Retrieval-augmented systems routinely present citations alongside generated answers, yet a citation does not confirm that the corresponding source mea

R2D-RL: A RoboCup 2D Soccer Environment for Multi-Agent Reinforcement Learning

Model ReleasesDGX agent

arXiv:2606.18786v2 Announce Type: replace Abstract: Robot soccer is a challenging testbed for multi-agent reinforcement learning because it combines partial observability, cooperative and adversarial

Reinforcement Learning without Ground-Truth Solutions can Improve LLMs

Model ReleasesDGX agent

arXiv:2606.27369v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) for training LLMs typically rely on ground-truth answers to assign rewards, limiting their applica

ResilPhase: Plug-and-Play Phase Mapping and Noise-Resilient Macro-Trajectory Extrapolation for Diffusion Acceleration

ResearchDGX agent

arXiv:2606.26769v1 Announce Type: new Abstract: The adoption of powerful diffusion models is hindered by their significant inference latency. Recent ``cache-then-forecast'' schemes alleviate this issu

Revisiting the Platonic Representation Hypothesis: An Aristotelian View

ResearchDGX agent

arXiv:2602.14486v2 Announce Type: replace-cross Abstract: The Platonic Representation Hypothesis suggests that representations from neural networks are converging to a common statistical model of real

Rolling Shutter Relative Pose Estimation Made Practical

Model ReleasesDGX agent

arXiv:2606.26863v1 Announce Type: new Abstract: Rolling shutter (RS) cameras equip virtually all consumer devices, yet RS-aware relative pose estimation has remained impractical: the state-of-the-art

Sakana Fugu Technical Report https://arxiv.org/abs/2606.21228 🐡

AgentsDGX agent

The Sakana Fugu Technical Report presents research on fugu (pufferfish), likely exploring novel approaches to AI model compression, efficiency, or optimization techniques, given Sakana AI's focus on e

Scalable AI-assisted Workflow Management for Detector Design Optimization Using Distributed Computing

Model ReleasesDGX agent

arXiv:2603.30014v2 Announce Type: replace-cross Abstract: The Production and Distributed Analysis (PanDA) system, originally developed for the ATLAS experiment at the CERN Large Hadron Collider (LHC),

SignSparK: Efficient Multilingual Sign Language Production via Sparse Keyframe Learning

ApplicationsDGX agent

arXiv:2603.10446v4 Announce Type: replace Abstract: Sign Language Production (SLP) faces a fundamental trade-off: direct text-to-pose models suffer from regression-to-the-mean effects, while dictionar

Sources: DeepSeek's $7.4B raise was prompted by the release of Mythos as CEO Liang Wenfeng realized DeepSeek couldn't compete without a massive war chest (The Information)

Model ReleasesDGX agent

The Information: Sources: DeepSeek's $7.4B raise was prompted by the release of Mythos as CEO Liang Wenfeng realized DeepSeek couldn't compete without a massive war chest — Up until two months ago, De

← Previous
1…639640641642643…1053
Next →