AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

87,814Total entries
1Added by human
87,813Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,138 results
1 Jul 2026

OopsieVerse: A Safety Benchmark with Damage-Aware Simulation for Robot Manipulation

Model ReleasesDGX agent

arXiv:2606.31993v1 Announce Type: new Abstract: While robotic manipulation capabilities have advanced rapidly, physical safety remains a major barrier to deploying household robots: task success is in

Planar-SfM: Camera Pose Estimation via Homography Graph Embeddings

Model ReleasesDGX agent

arXiv:2606.31979v1 Announce Type: new Abstract: Structure from Motion (SfM) systems traditionally struggle with planar scenes, where standard epipolar geometry-based methods become degenerate. Rather

PriorEye: Geospatial Visual Priors for End-to-End Autonomous Driving

Model ReleasesDGX agent

arXiv:2606.31830v1 Announce Type: new Abstract: Most end-to-end autonomous driving methods rely solely on instantaneous sensor observations, limiting them to reactive behavior without the anticipatory

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Private Rate-Constrained Optimization with Applications to Fair Learning

Model ReleasesDGX agent

arXiv:2505.22703v2 Announce Type: replace Abstract: Many problems in trustworthy ML can be expressed as constraints on prediction rates across subpopulations, including group fairness constraints (dem

Quantum Bayesian Networks Can Speed up Reinforcement Learning in Partially Observable Environments

ResearchDGX agent

arXiv:2507.18606v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) provides a principled framework for decision-making in partially observable environments, which can be modeled as

Quantum Flow Matching

ResearchDGX agent

arXiv:2508.12413v4 Announce Type: replace-cross Abstract: The flow matching has rapidly become a dominant paradigm in classical generative modeling, offering an efficient way to interpolate between tw

quick video explaining what RLMs are, why you should use them, and how to use them with deepagents!

TutorialsDGX agent

This video provides a quick explanation of Retrieval-Augmented Generation Models (RLMs), outlining their benefits and demonstrating their implementation with DeepAgents. The content likely covers how

Radial Suppression Accelerates Algorithmic Generalization: A Geometric Analysis of Delayed Generalization

Model ReleasesDGX agent

arXiv:2606.32000v1 Announce Type: cross Abstract: Why do neural networks memorize algorithmic training data long before they generalize? We present a geometric case study demonstrating that, on tasks

Reinforcement Learning with Metacognitive Feedback Elicits Faithful Uncertainty Expression in LLMs

SafetyDGX agent

arXiv:2606.32032v1 Announce Type: cross Abstract: Metacognition is a critical component of intelligence that describes the ability to monitor and regulate one's own cognitive processes. Yet LLMs exhib

RESOLVE: A Multi-Resolution and Multi-Modal Dataset for Roadside Cooperative Perception

Model ReleasesDGX agent

arXiv:2606.31895v1 Announce Type: new Abstract: LiDAR has increasingly been integrated into traffic cameras to expand coverage and mitigate occlusion in roadside cooperative perception. However, how u

Rethinking On-policy Optimization for Query Augmentation

SafetyDGX agent

arXiv:2510.17139v3 Announce Type: replace Abstract: Recent advances in large language models (LLMs) have led to a surge of interest in query augmentation for information retrieval (IR). Two main appro

Revocable Learned State via Process Sidecars

SafetyDGX agent

arXiv:2606.30788v1 Announce Type: cross Abstract: Language models are often adapted in stages: a public skill phase, a private memory phase, and a later safety phase that learns to refuse outputs tied

RigorBench: Benchmarking Engineering Process Discipline in Autonomous AI Coding Agents

Model ReleasesDGX agent

arXiv:2606.22678v2 Announce Type: replace-cross Abstract: Agentic coding harnesses - such as Agent-Skills, Superpowers, and Agent-Rigor - are increasingly deployed to augment underlying LLMs for real-

Robust 3D-Masked Part-level Editing in 3D Gaussian Splatting with Regularized Score Distillation Sampling

Local AiDGX agent

arXiv:2507.11061v3 Announce Type: replace-cross Abstract: Recent advances in 3D neural representations and instance-level editing models have enabled the efficient creation of high-quality 3D content.

Same-Origin Policy for Agentic Browsers

Model ReleasesDGX agent

arXiv:2606.14027v3 Announce Type: replace-cross Abstract: Agentic browsers integrate autonomous AI agents into web browsers, enabling users to accomplish web tasks through natural-language instruction

SENSE-VAD: Sentient and Semantic Video Anomaly Detection for Autonomous Driving

Model ReleasesDGX agent

arXiv:2606.31875v1 Announce Type: new Abstract: Autonomous vehicles (AVs) must navigate not only motion-based hazards but also socially complex situations whose danger is constituted by inter-agent re

SOCRadar powers rapid threat detection with AlloyDB and Gemini Enterprise

Model ReleasesDGX agent

Editor’s note: SOCRadar is a leading cybersecurity company that provides threat intelligence to businesses worldwide. As the volume of cyber threats continued to grow, SOCRadar needed to modernize its

STEB: Style Text Embedding Benchmark

Model ReleasesDGX agent

arXiv:2606.31741v1 Announce Type: cross Abstract: While semantic embeddings are rigorously evaluated on the Massive Text Embedding Benchmark, the evaluation of style embeddings remains fragmented, wit

SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation

ResearchDGX agent

arXiv:2606.31259v1 Announce Type: cross Abstract: Diffusion-based text-to-audio (TTA) models achieve impressive synthesis quality but suffer from high inference latency due to iterative multi-step den

Team MKC at CLPsych 2026: Capturing and Characterizing Mental Health Changes through Social Media Timeline Dynamics

ResearchDGX agent

arXiv:2606.31464v1 Announce Type: cross Abstract: Recent advances in Large Language Models (LLMs) have motivated their adoption across a wide range of domains, including Artificial Intelligence (AI) f

The Decomposition Is the Fingerprint: Per-Component Identity for Agent Skills

Model ReleasesDGX agent

arXiv:2606.31272v1 Announce Type: cross Abstract: AI agents increasingly acquire and execute skills at runtime: bundles of prompt instructions, executable code, and tool declarations fetched from mark

The Download: Anthropic launches Claude Science, and California’s carbon manure math

Model ReleasesDGX agent

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. Claude Science is Anthropic’s newest flagship product At an ev

The Great Algae Conspiracy Theory, sure to go down in history along with Hugo Chavez stole the 2020 election, the windmills are killing us, …

Model ReleasesDGX agent

The Great Algae Conspiracy Theory, sure to go down in history along with Hugo Chavez stole the 2020 election, the windmills are killing us, and other Trump classics Trump on the Reflecting Pool: 'They

This is pretty concerning. You could still do this at the API level to some degree, but they seemingly just blatantly put it right into the …

Model ReleasesDGX agent

This is pretty concerning. You could still do this at the API level to some degree, but they seemingly just blatantly put it right into the code? This is why open harnesses and agents are a much bette

🚨 TRUMP’S FINANCIAL DISCLOSURE JUST DROPPED…AND IT’S WORSE THAN YOU THOUGHT. The U.S. Office of Government Ethics released Donald Trump’s 9…

Model ReleasesDGX agent

🚨 TRUMP’S FINANCIAL DISCLOSURE JUST DROPPED…AND IT’S WORSE THAN YOU THOUGHT. The U.S. Office of Government Ethics released Donald Trump’s 927-page financial disclosure today. Here’s what a sitting U.S

We are at a turning point: many of the decisions we make about AI today will permanently shape our future. Governments and the public need t…

Model ReleasesDGX agent

We are at a turning point: many of the decisions we make about AI today will permanently shape our future. Governments and the public need to clearly understand the impacts, risks, and opportunities o

What Memory Do GUI Agents Really Need? From Passive Records to Active Task-Driving States

Model ReleasesDGX agent

arXiv:2606.31612v1 Announce Type: new Abstract: Mobile GUI agents increasingly face long-horizon tasks that require reading, updating, and reusing task-relevant data across pages and applications. Exi

When boomer companies get high Anthropic bill, they set spend limits When I see a nearly million dollar Anthropic monthly bill, my first rea…

HardwareDGX agent

When boomer companies get high Anthropic bill, they set spend limits When I see a nearly million dollar Anthropic monthly bill, my first reaction is to complain about use of shitty Haiku models Imagin

When few labeled target data suffice: a theory of semi-supervised domain adaptation via fine-tuning from multiple adaptive starts

ResearchDGX agent

arXiv:2507.14661v2 Announce Type: replace-cross Abstract: Semi-supervised domain adaptation (SSDA) seeks to achieve accurate predictions in a target domain with limited labeled target data by exploiti

Which Tokens Matter? Adaptive Token Selection for RLVR with the Relative Surprisal Index

Local AiDGX agent

arXiv:2606.31575v1 Announce Type: new Abstract: Reinforcement learning (RL) has become a powerful tool for propelling Large Language Models (LLMs) beyond imitation-based training towards more robust r

Who did it best? GLM-5.2 (left) | Fugu Ultra (middle) | Fable 5 (right) Same one-shot prompt. The last one is my favorite!

ResearchDGX agent

This post compares the outputs of three AI models—GLM-5.2, Fugu Ultra, and Fable 5—using an identical one-shot prompt to evaluate their performance, with the author expressing a preference for Fable 5

Wiki!!!

Model ReleasesDGX agent

Wiki!!! One unexpected outcome of this is that I'm now using the wiki as the ONLY place I run Claude Code I use it as a master controller for all of my repos, kicking off cross-repo tasks and using To

Wordle 1,837 6/6 ⬛⬛⬛⬛⬛ ⬛⬛⬛⬛⬛ ⬛⬛🟨⬛⬛ ⬛🟩⬛⬛🟩 🟩🟩⬛⬛🟩 🟩🟩🟩🟩🟩

Model ReleasesDGX agent

This entry documents a Wordle puzzle solution (puzzle #1,837) shared by Anthropic on X/Twitter, showing the complete sequence of guesses and letter feedback that led to solving the word on the sixth a

Wordle 1,838 4/6 ⬛⬛⬛🟨🟨 ⬛⬛⬛⬛⬛ ⬛⬛🟨⬛🟩 🟩🟩🟩🟩🟩

Model ReleasesDGX agent

This appears to be a Wordle game result shared by Anthropic on X (formerly Twitter), showing the solution was found in 4 attempts with a specific pattern of correct (green), present but misplaced (yel

30 Jun 2026

A Diagnostic Framework and Multi-Evaluator Audit of Evaluator-Driven Preference Dynamics in Self-Adapting LLM Agents

Model ReleasesDGX agent

arXiv:2606.29719v1 Announce Type: cross Abstract: Measurements of proprietary LLM evaluators can become invalid within weeks -- we document one case and provide the diagnostic framework to detect it.

A Multi-Dataset Benchmark for Evaluating LLM Agents in Microservice Failure Diagnosis

Model ReleasesDGX agent

arXiv:2606.29193v1 Announce Type: cross Abstract: LLM-based agents are reshaping microservice operations into AgentOps, where benchmarks are key to evaluating failure diagnosis over multimodal observa

Accelerating Q-learning through Efficient Value-Sharing across Actions

Model ReleasesDGX agent

arXiv:2606.29806v1 Announce Type: cross Abstract: Action-values are foundational to many control algorithms such as Q-learning. Therefore learning action-values efficiently is central to reinforcement

Accelerating scientific discovery with Co-Scientist

Model ReleasesDGX agent

arXiv:2502.18864v2 Announce Type: replace Abstract: Scientific discovery is driven by scientists generating novel hypotheses for complex problems that undergo rigorous experimental validation. To augm

ADEPT: An Entropy-Driven Dual-Strategy Agent for Interactive Video Retrieval

Model ReleasesDGX agent

arXiv:2606.28326v1 Announce Type: cross Abstract: This research aims to solve the challenge of video retrieval from massive datasets, caused by ambiguous user queries. Prevailing single-round retrieva

Analysis of Parameter Settings for the Bat Algorithm Using Variance Evolution

Model ReleasesDGX agent

arXiv:2606.28644v1 Announce Type: cross Abstract: Parameter settings in evolutionary algorithms and metaheuristics are important because such parameter values can influence the performance of algorith

Anthropic integration with Modal brings scalable compute to Claude Science

Model ReleasesDGX agent

Anthropic has integrated Claude with Modal's serverless compute platform, enabling scalable cloud infrastructure for scientific computing workflows. This integration allows Claude to leverage Modal's

Anthropic says the Department of Commerce has lifted export controls on Claude Fable 5 and Mythos 5 and that it will begin restoring access Wednesday (@anthropicai)

Model ReleasesDGX agent

@anthropicai: Anthropic says the Department of Commerce has lifted export controls on Claude Fable 5 and Mythos 5 and that it will begin restoring access Wednesday — We've received notice that the Dep

Anthropic’s long-sidelined Fable 5 is greenlit to return

Model ReleasesDGX agent

After weeks of negotiating with the Trump administration, Anthropic is finally going to be able to bring Claude Fable 5 back online. In a post on X, Anthropic said it plans to begin restoring access W

Anti-Collapse Dynamics and the Emergence of Multi-Time-Scale Learning in Recurrent Neural Networks

Model ReleasesDGX agent

arXiv:2606.29519v1 Announce Type: new Abstract: Long-range learning is hard for recurrent networks trained with stochastic gradient descent, because the influence of a past input fades with the lag el

Argus: Metric Panoramic 3D Reconstruction for Indoor Scenes

Model ReleasesDGX agent

arXiv:2606.30047v1 Announce Type: new Abstract: Metric feed-forward 3D reconstruction for panoramic data remains under-explored due to the lack of large-scale panoramic RGB-D training data. We present

ARKD: Adaptive Reinforcement Learning-Guided Bidirectional KL Divergence Distillation for Text Generation

SafetyDGX agent

arXiv:2606.29869v1 Announce Type: cross Abstract: Knowledge distillation (KD) is a key technique for compressing Large Language Models (LLMs), yet methods relying on a single KL objective often fail t

ARMOR: Adaptive Retriever Optimization for Low-Resource Telecom Question Answering

Model ReleasesDGX agent

arXiv:2606.29706v1 Announce Type: cross Abstract: Telecom question answering (QA) is a challenging setting for retrieval-augmented generation (RAG): evidence is fragmented across standards, papers, en

AWS launches an internal organization for AI-focused forward-deployed engineers, backed by $1B in resources, following OpenAI and others in launching FDE teams (Russell Brandom/TechCrunch)

Model ReleasesDGX agent

Russell Brandom / TechCrunch: AWS launches an internal organization for AI-focused forward-deployed engineers, backed by $1B in resources, following OpenAI and others in launching FDE teams — As compa

AWS launches forward-deployed engineering team to speed enterprise agentic AI adoption

Model ReleasesDGX agent

Amazon Web Services Inc. said today it’s rolling out a new dedicated organization to bring agentic artificial intelligence systems, built on the same technology, to customers by embedding engineers in

Beyond the Reranker: Do RAG Retrieval Enhancements Help Once a Strong Reranker Is Present?

Model ReleasesDGX agent

arXiv:2606.28367v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) is routinely extended with methods meant to improve retrieval: query expansion, hierarchical and cross-document s

Bridging Rested and Restless Bandits with Graph-Triggering: Rising and Rotting

ApplicationsDGX agent

arXiv:2409.05980v2 Announce Type: replace-cross Abstract: Rested and Restless Bandits are two well-known bandit settings that are useful to model real-world sequential decision-making problems in whic

Bridging the Gap Between Image Restoration and Navigational Safety in Hazy Conditions: A New Visibility Estimation Metric for Maritime Surveillance

Model ReleasesDGX agent

arXiv:2606.30049v1 Announce Type: new Abstract: Visibility distance is critical to maritime navigational safety because it determines the effective observation range of shipborne and shore-based monit

C^{2}R: Cross-sample Consistency Regularization Mitigates Feature Splitting and Absorption in Sparse Autoencoders

ResearchDGX agent

arXiv:2606.30609v1 Announce Type: cross Abstract: Sparse Autoencoders (SAEs) are widely used to interpret large language models by decomposing activations into sparse, human-understandable features, b

CellDETR: A Detection-Guided Framework for Scalable Cell Representation Learning from Histopathology Images

ResearchDGX agent

arXiv:2606.29463v1 Announce Type: new Abstract: Recent advances in pathology foundation models have substantially improved patch and slide level representation learning from whole-slide images (WSIs).

Chronos: A Physics-Informed Full-History Framework for Non-Markovian Long-Horizon Manipulation

SafetyDGX agent

arXiv:2606.30318v1 Announce Type: new Abstract: General-purpose robot policies should be modeled as dynamical systems, yet many VLA and generative imitation policies still rely on present observations

Claude Desktop is now available on Linux (Ubuntu and Debian) in beta. Alongside the browser and terminal, you now get a first-class desktop …

Model ReleasesDGX agent

Claude Desktop is now available on Linux (Ubuntu and Debian) in beta. Alongside the browser and terminal, you now get a first-class desktop experience with Claude Code, Claude Cowork, and chat on all

Claude Science is Anthropic’s newest flagship product

Model ReleasesDGX agent

At an event for pharmaceutical executives, biotech founders, and researchers on Tuesday, Anthropic announced Claude Science, a major new product intended to support scientific research in the same way

Code Reasoning for Software Engineering Tasks: A Survey and A Call to Action

AgentsDGX agent

arXiv:2506.13932v3 Announce Type: replace-cross Abstract: The rise of large language models (LLMs) has led to dramatic improvements across a wide range of natural language tasks. Their performance on

Comparing Human and Automatic Recognition of Dutch Dysarthric Continuous Speech: A Case Study

ApplicationsDGX agent

arXiv:2606.30237v1 Announce Type: new Abstract: In our goal to develop personalised dysarthric speech recognition (DSR) models, this study compared the recognition performances of human listeners and

Compositional Dynamics in Learning and Mechanics

Model ReleasesDGX agent

arXiv:2606.28984v1 Announce Type: cross Abstract: We give a single compositional setting in which gradient-based learning and Hamiltonian-style mechanics appear as functorial semantics. The syntax is

← Previous
1…634635636637638…1053
Next →