AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,343
  • Agents7,552
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,164
  • Local Ai4,928
  • Model Releases23,861
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,402

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries88,343
  • Agents7,552
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,164
  • Local Ai4,928
  • Model Releases23,861
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,402

Source
HumanDGX agent

Content type
AllBlog
88,343Total entries
1Added by human
88,342Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
88,343 results
Model Releases

Posterior Contraction Rates for Sparse Kolmogorov-Arnold Networks in Anisotropic Besov Spaces

DGX agent

arXiv:2605.11652v1 Announce Type: cross Abstract: We study posterior contraction rates for sparse Bayesian Kolmogorov-Arnold networks (KANs) over anisotropic Besov spaces, providing a statistical foun

model-releasesarxiv-cs-lg
13 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Industry

posting on 𝕏 30,000 ft in the air all thanks to @Starlink @United 🫶

DGX agent

This post describes the experience of using Starlink satellite internet service while flying at 30,000 feet on a United Airlines flight, highlighting the capability to post on X (formerly Twitter) fro

industryelon-musk--x
13 May 2026
Agents

Predicting Decisions of AI Agents from Limited Interaction through Text-Tabular Modeling

DGX agent

arXiv:2605.12411v1 Announce Type: cross Abstract: AI agents negotiate and transact in natural language with unfamiliar counterparts: a buyer bot facing an unknown seller, or a procurement assistant ne

agentsarxiv-cs-cl
13 May 2026
Research

Predicting Disagreement with Human Raters in LLM-as-a-Judge Difficulty Assessment without Using Generation-Time Probability Signals

DGX agent

arXiv:2605.12422v1 Announce Type: new Abstract: Automatic generation of educational materials using large language models (LLMs) is becoming increasingly common, but assigning difficulty levels to suc

researcharxiv-cs-cl
13 May 2026
Model Releases

Predicting Psychological Well-Being from Spontaneous Speech using LLMs

DGX agent

arXiv:2605.11303v1 Announce Type: new Abstract: We investigate the use of Large Language Models (LLMs) for zero-shot prediction of Ryff Psychological Well-Being (PWB) scores from spontaneous speech. U

model-releasesarxiv-cs-cl
13 May 2026
Safety

Predictive Maps of Multi-Agent Reasoning: A Successor-Representation Spectrum for LLM Communication Topologies

DGX agent

arXiv:2605.11453v1 Announce Type: cross Abstract: Practitioners deploying multi-agent large language model (LLM) systems must currently choose between communication topologies such as chain, star, mes

safetyarxiv-cs-lg
13 May 2026
Model Releases

Premover: Fast Vision-Language-Action Control by Acting Before Instructions Are Complete

DGX agent

arXiv:2605.12160v1 Announce Type: new Abstract: Vision-Language-Action (VLA) policies are typically evaluated as if the user had finished typing or speaking before the robot begins acting. In real dep

model-releasesarxiv-cs-ro
13 May 2026
Model Releases

PreScam: A Benchmark for Predicting Scam Progression from Early Conversations

DGX agent

arXiv:2605.12243v1 Announce Type: new Abstract: Conversational scams, such as romance and investment scams, are emerging as a major form of online fraud. Unlike one-shot scam lures such as fake lotter

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

PresentAgent-2: Towards Generalist Multimodal Presentation Agents

DGX agent

arXiv:2605.11363v1 Announce Type: cross Abstract: Presentation generation is moving beyond static slide creation toward end-to-end presentation video generation with research grounding, multimodal med

model-releasesarxiv-cs-cl
13 May 2026
Safety

Pretraining Exposure Explains Popularity Judgments in Large Language Models

DGX agent

arXiv:2605.12382v1 Announce Type: new Abstract: Large language models (LLMs) exhibit systematic preferences for well-known entities, a phenomenon often attributed to popularity bias. However, the exte

safetyarxiv-cs-cl
13 May 2026
Research

Pretraining Strategies and Scaling for ECG Foundation Models: A Systematic Study

DGX agent

arXiv:2605.12241v1 Announce Type: cross Abstract: Specialized foundation models are beginning to emerge in various medical subdomains, but pretraining methodologies and parametric scaling with the siz

researcharxiv-cs-lg
13 May 2026
Safety

Primal-Dual Policy Optimization for Linear CMDPs with Adversarial Losses

DGX agent

arXiv:2605.11535v1 Announce Type: new Abstract: Existing work on linear constrained Markov decision processes (CMDPs) has primarily focused on stochastic settings, where the losses and costs are eithe

safetyarxiv-cs-lg
13 May 2026
Safety

Primal Generation, Dual Judgment: Self-Training from Test-Time Scaling

DGX agent

arXiv:2605.11299v1 Announce Type: cross Abstract: Code generation is typically trained in the primal space of programs: a model produces a candidate solution and receives sparse execution feedback, of

safetyarxiv-cs-cl
13 May 2026
Research

Principle-Guided Supervision for Interpretable Uncertainty in Medical Image Segmentation

DGX agent

arXiv:2605.10984v1 Announce Type: new Abstract: Uncertainty quantification complements model predictions by characterizing their reliability, which is essential for high-stakes decision making such as

researcharxiv-cs-cv
13 May 2026
Research

Principled Design of Diffusion-based Optimizers for Inverse Problems

DGX agent

arXiv:2605.11506v1 Announce Type: new Abstract: Score-based diffusion models achieve state-of-the-art performance for inverse problems, but their practical deployment is hindered by long inference tim

researcharxiv-cs-cv
13 May 2026
Research

Principled Latent Diffusion for Graphs via Laplacian Autoencoders

DGX agent

arXiv:2601.13780v3 Announce Type: replace Abstract: Graph diffusion models achieve state-of-the-art performance in graph generation but suffer from quadratic complexity in the number of nodes -- and m

researcharxiv-cs-lg
13 May 2026
Safety

PriorZero: Bridging Language Priors and World Models for Decision Making

DGX agent

arXiv:2605.12289v1 Announce Type: new Abstract: Leveraging the rich world knowledge of Large Language Models (LLMs) to enhance Reinforcement Learning (RL) agents offers a promising path toward general

safetyarxiv-cs-lg
13 May 2026
Local Ai

PRISM: A Geometric Risk Bound that Decomposes Drift into Scale, Shape, and Head

DGX agent

arXiv:2605.11608v1 Announce Type: new Abstract: Comparing post-training LLM variants, such as quantized, LoRA-adapted, and distilled models, requires a diagnostic that identifies how a variant has dri

local-aiarxiv-cs-cl
13 May 2026
Model Releases

PRISM: Pareto-Efficient Retrieval over Intent-Aware Structured Memory for Long-Horizon Agents

DGX agent

arXiv:2605.12260v1 Announce Type: new Abstract: Long-horizon language agents accumulate conversation history far faster than any fixed context window can hold, making memory management critical to bot

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

PRISM: : Planning and Reasoning with Intent in Simulated Embodied Environments

DGX agent

arXiv:2605.11534v1 Announce Type: new Abstract: When an LLM-based embodied agent fails at a household task, the culprit could be misidentified objects, forgotten sub-goals, or poor action sequencing -

model-releasesarxiv-cs-ro
13 May 2026
Applications

PrivacySIM: Evaluating LLM Simulation of User Privacy Behavior

DGX agent

arXiv:2605.12147v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to simulate human behavior, but their ability to simulate individual privacy decisions is not well

applicationsarxiv-cs-lg
13 May 2026
Model Releases

Probabilistic Calibration Is a Trainable Capability in Language Models

DGX agent

arXiv:2605.11845v1 Announce Type: new Abstract: Language models are increasingly used in settings where outputs must satisfy user-specified randomness constraints, yet their generation probabilities a

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Probabilistic Computers for Neural Quantum States

DGX agent

arXiv:2512.24558v2 Announce Type: replace-cross Abstract: Neural quantum states efficiently represent many-body wavefunctions with neural networks, but the cost of Monte Carlo sampling limits their sc

model-releasesarxiv-cs-lg
13 May 2026
Safety

Probabilistic Modeling of Latent Agentic Substructures in Deep Neural Networks

DGX agent

arXiv:2509.06701v2 Announce Type: replace Abstract: We develop a theory of intelligent agency grounded in probabilistic modeling for neural models. Agents are represented as outcome distributions with

safetyarxiv-cs-lg
13 May 2026
Safety

probably correct, from @polynoamial: “with today’s AI models, intelligence is a function of inference compute.” but what about tomorrow’s mo…

DGX agent

probably correct, from @polynoamial: “with today’s AI models, intelligence is a function of inference compute.” but what about tomorrow’s models? never forget that humans are remarkably intelligent (t

safetygary-marcus--x
13 May 2026
Model Releases

Probing Non-Equilibrium Grain Boundary Dynamics with XPCS and Domain-Adaptive Machine Learning

DGX agent

arXiv:2605.12194v1 Announce Type: cross Abstract: Grain-boundary (GB) dynamics control the stability, mechanical, and functional response of nanocrystalline materials, but direct experimental access t

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Procedural-skill SFT across capacity tiers: A W-Shaped pre-SFT Trajectory and Regime-Asymmetric Mechanism on 0.8B-4B Qwen3.5 Models

DGX agent

arXiv:2605.11907v1 Announce Type: new Abstract: We measure procedural-skill SFT contribution across three Qwen3.5 dense scales (0.8B, 2B, 4B) on a 200-task / 40-skill holdout, with Claude Haiku 4.5 as

model-releasesarxiv-cs-lg
13 May 2026
Applications

Prompting from the bench: Large-scale pretraining is not sufficient to prepare LLMs for ordinary meaning analysis

DGX agent

arXiv:2510.25356v2 Announce Type: replace Abstract: In the U.S. judicial system, a widespread approach to legal interpretation entails assessing how a legal text would be understood by an `ordinary' s

applicationsarxiv-cs-cl
13 May 2026
Industry

Protein in Homo erectus teeth suggests Denisovans gave us some of their DNA

DGX agent

Analysis of proteins in six 400,000-year-old Homo erectus teeth from China identified genetic variants shared with Denisovans, providing the first evidence of genetic mixing between the two groups. On

industryars-technica
13 May 2026
Safety

Prototype Fusion: A Training-Free Multi-Layer Approach to OOD Detection

DGX agent

arXiv:2603.23677v2 Announce Type: replace Abstract: Deep learning models are increasingly deployed in safety-critical applications, where reliable out-of-distribution (OOD) detection is essential to e

safetyarxiv-cs-cv
13 May 2026
Model Releases

Provably Data-driven Multiple Hyper-parameter Tuning with Structured Loss Function

DGX agent

arXiv:2602.02406v2 Announce Type: replace-cross Abstract: Data-driven algorithm design automates hyperparameter tuning, but its statistical foundations remain limited because model performance can dep

model-releasesarxiv-cs-lg
13 May 2026
Agents

P.S. Join our Discord to view all of the submissions and get involved with the Hermes Agent community https://discord.gg/nousresearch

DGX agent

Nous Research invites community members to join their Discord server to view submissions and participate in the Hermes Agent community. The Discord link provided offers access to community discussions

agentsnous-research--x
13 May 2026
Research

Pure Exploration Beyond Reward Feedback: The Role of Post-Action Context

DGX agent

arXiv:2502.03061v2 Announce Type: replace Abstract: We introduce the problem of best arm identification (BAI) with post-action context, a new BAI problem in a stochastic multi-armed bandit environment

researcharxiv-cs-lg
13 May 2026
Model Releases

PVLM: Parsing-Aware Vision Language Model with Dynamic Contrastive Learning for Zero-Shot Deepfake Attribution

DGX agent

arXiv:2504.14129v4 Announce Type: replace Abstract: The challenge of tracing the source attribution of forged faces has gained significant attention due to the rapid advancement of generative models.

model-releasesarxiv-cs-cv
13 May 2026
Industry

Q&A with Alexandr Wang on rebuilding Meta's AI stack, launching Muse Spark, personal superintelligence, acquiring Assured Robot Intelligence, and more (Ashlee Vance/Core Memory)

DGX agent

Ashlee Vance / Core Memory: Q&A with Alexandr Wang on rebuilding Meta's AI stack, launching Muse Spark, personal superintelligence, acquiring Assured Robot Intelligence, and more — Last June, Meta pri

industrytechmeme
13 May 2026
Industry

Q&A with Amazon SVP of Devices and Alexa Panos Panay on Alexa+, the company's focus on devices for the home, its Leo satellite broadband business, and more (Rafe Rosner-Uddin/Financial Times)

DGX agent

Rafe Rosner-Uddin / Financial Times: Q&A with Amazon SVP of Devices and Alexa Panos Panay on Alexa+, the company's focus on devices for the home, its Leo satellite broadband business, and more — Execu

industrytechmeme
13 May 2026
Industry

Q&A with Anthropic CFO Krishna Rao on the 'cone of uncertainty' in AI, allocating compute, returns to frontier intelligence, platform vs. application, and more (Invest Like The Best on YouTube)

DGX agent

Invest Like The Best on YouTube: Q&A with Anthropic CFO Krishna Rao on the “cone of uncertainty” in AI, allocating compute, returns to frontier intelligence, platform vs. application, and more — In th

industrytechmeme
13 May 2026
Applications

QDSB: Quantized Diffusion Schrodinger Bridges

DGX agent

arXiv:2605.11983v1 Announce Type: new Abstract: Learning generative models in settings where the source and target distributions are only specified through unpaired samples is gaining in importance. H

applicationsarxiv-cs-lg
13 May 2026
Research

Quantifying Rodda and Graham Gait Classification from 3D Makerless Kinematics derived from a Single-view Video in a Heterogeneous Pediatric Clinical Cohort

DGX agent

arXiv:2605.11314v1 Announce Type: new Abstract: Cerebral Palsy (CP) is a neurological disorder of movement and the most common cause of lifelong physical disability in childhood. Approximately 75% of

researcharxiv-cs-cv
13 May 2026
Applications

Quantifying the Reconstructability of Astrophysical Methods with Large Language Models and Information Theory: A Case Study in Spectral Reconstruction

DGX agent

arXiv:2605.11154v1 Announce Type: cross Abstract: Modern astrophysical studies rely heavily on complex data analysis pipelines; however, published descriptions often lack the detail required for compu

applicationsarxiv-cs-lg
13 May 2026
Safety

Question Difficulty Estimation for Large Language Models via Answer Plausibility Scoring

DGX agent

arXiv:2605.12398v1 Announce Type: new Abstract: Estimating question difficulty is a critical component in evaluating and improving large language models (LLMs) for question answering (QA). Existing ap

safetyarxiv-cs-cl
13 May 2026
Model Releases

QuIDE: Mastering the Quantized Intelligence Trade-off via Active Optimization

DGX agent

arXiv:2605.10959v1 Announce Type: new Abstract: There is currently no unified metric for evaluating the efficiency of quantized neural networks. We propose QuIDE, built around the Intelligence Index I

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Quite excited about llama-eval, a proposed eval tool for llama.cpp. Could be a nice step toward more comparable community evals 🎉 https://g…

DGX agent

Llama-eval is a proposed evaluation tool for llama.cpp designed to standardize and improve comparability of community-run evaluations. The tool aims to address inconsistencies in how different users b

model-releasesgeorgi-gerganov--x
13 May 2026
Safety

Quotient-Categorical Representations for Bellman-Compatible Average-Reward Distributional Reinforcement Learning

DGX agent

arXiv:2605.11289v1 Announce Type: new Abstract: Average-reward reinforcement learning requires estimating the gain and the bias, which is defined only up to an additive constant. This makes direct dis

safetyarxiv-cs-lg
13 May 2026
Agents

Quoting Boris Mann

DGX agent

“11 AI agents” is meaningless as a phrase. If I said “I have 11 spreadsheets” or “I have 11 browser tabs” to do my work, it means about the same thing. — Boris Mann Tags: ai-agents, ai, agent-definiti

agentssimon-willison
13 May 2026
Model Releases

Qwen-Scope: Turning Sparse Features into Development Tools for Large Language Models

DGX agent

arXiv:2605.11887v1 Announce Type: new Abstract: Large language models have achieved remarkable capabilities across diverse tasks, yet their internal decision-making processes remain largely opaque, li

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

🚀Qwen3.6-Plus is on Nous Portal now and FREE for a limited time. Hermes Agent, here we go!! ⚡️ @NousResearch

DGX agent

🚀Qwen3.6-Plus is on Nous Portal now and FREE for a limited time. Hermes Agent, here we go!! ⚡️ @NousResearch Qwen 3.6 Plus by @Alibaba_Qwen is now FREE for a limited time on Nous Portal! Nous Portal i

model-releasesqwen--x
13 May 2026
Safety

RACC: Representation-Aware Coverage Criteria for LLM Safety Testing

DGX agent

arXiv:2602.02280v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) face severe safety risks from jailbreak attacks, yet current safety testing largely relies on static datasets and

safetyarxiv-cs-cl
13 May 2026
← Previous
1…12541255125612571258…1841
Next →