AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,724 results
Research

AI4SLT: Empirical Processes in Lean 4 for Formal Statistical Learning Theory

DGX agent

arXiv:2602.02285v2 Announce Type: replace-cross Abstract: We present the first comprehensive Lean 4 formalization of statistical learning theory (SLT) grounded in empirical process theory. Our en-to-e

researcharxiv-cs-cl
11 Jun 2026
Safety
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Automating Geometry-Intensive Compliance Checking in BIM: Graph-Based Semantic Reasoning Framework

DGX agent

arXiv:2606.12065v1 Announce Type: new Abstract: Automating compliance check for geometry-intensive regulations remains a significant technical bottleneck in Building Information Modeling (BIM), primar

safetyarxiv-cs-ai
11 Jun 2026
Research

EvoLMM: Self-Evolving Large Multimodal Models with Continuous Rewards

DGX agent

arXiv:2511.16672v4 Announce Type: replace Abstract: Recent advances in large multimodal models (LMMs) have enabled impressive reasoning and perception abilities, yet most existing training pipelines s

researcharxiv-cs-cv
11 Jun 2026
Safety

Implicit Neural Representations of Individual Behavior

DGX agent

arXiv:2606.12200v1 Announce Type: cross Abstract: We study policy representation learning from unlabeled multi-policy behavioral data. Each episode is generated by a fixed policy, but policy labels ar

safetyarxiv-cs-ai
11 Jun 2026
Model Releases

MobileFineTuner: A Mobile-Native Framework for On-Device LLM Fine-Tuning in Real-World Embedded AI Applications

DGX agent

arXiv:2512.08211v2 Announce Type: replace Abstract: Large language models (LLMs) are moving from cloud-centric services toward on-device embedded AI, where models interact with private, longitudinal s

model-releasesarxiv-cs-lg
11 Jun 2026
Safety

Reinforcement Learning with Action-Triggered Observations

DGX agent

arXiv:2510.02149v2 Announce Type: replace Abstract: We introduce Action-Triggered Sporadically Traceable Markov Decision Processes (ATST-MDPs), a reinforcement learning framework for partial observabi

safetyarxiv-cs-lg
11 Jun 2026
Safety

Sample-Efficient Hypergradient Estimation for Decentralized Bi-Level Reinforcement Learning

DGX agent

arXiv:2603.14867v4 Announce Type: replace-cross Abstract: Many strategic decision-making problems, such as environment design for warehouse robots, can be naturally formulated as bi-level reinforcemen

safetyarxiv-cs-ai
11 Jun 2026
Model Releases

TAHOE: Text-to-SQL with Automated Hint Optimization from Experience

DGX agent

arXiv:2606.12387v1 Announce Type: cross Abstract: Large Language Models (LLMs) have democratized database access through Text-to-SQL, but moving from prototypes to production remains difficult. Real d

model-releasesarxiv-cs-ai
11 Jun 2026
Safety

The Algorithm Is Not the Behavior: Learned Priors Override Look-Ahead in a Chess-Playing Neural Network

DGX agent

arXiv:2508.21380v3 Announce Type: replace-cross Abstract: Recent mechanistic work has uncovered learned algorithms within neural networks, from modular arithmetic to search and planning in game-playin

safetyarxiv-cs-ai
11 Jun 2026
Model Releases

When Does Language Matter? Multilingual Instructions Reveal Step-wise Language Sensitivity in Vision-Language-Action Models

DGX agent

arXiv:2606.11906v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have shown strong performance in language-conditioned robotic manipulation, yet their robustness to linguistic varia

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

When Generic Prompt Improvements Hurt: Evaluation-Driven Iteration for LLM Applications

DGX agent

arXiv:2601.22025v2 Announce Type: replace-cross Abstract: Evaluating Large Language Model (LLM) applications differs from conventional software testing because outputs are probabilistic, semantically

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Which Models Are Our Models Built On? Auditing Invisible Dependencies in Modern LLMs

DGX agent

arXiv:2606.12385v1 Announce Type: new Abstract: Modern LLM training pipelines increasingly rely on other models to generate data, filter corpora, judge outputs, and guide development decisions. These

model-releasesarxiv-cs-cl
11 Jun 2026
Industry

Build an AI-Powered Equipment Repair Assistant Using Amazon Bedrock AgentCore

DGX agent

In this post, you build an AI-powered equipment repair assistant using Amazon Bedrock AgentCore that helps farmers and field technicians diagnose equipment problems, identify required parts, and acces

industryaws-ml-blog
10 Jun 2026
Local Ai

Demo: Turn Research Into a Client-Ready Report with Row-Bot

DGX agent

Row-Bot is a local-first desktop AI assistant that orchestrates tools and models to handle reasoning and workflows while keeping data local. The demo likely showcases how Row-Bot's integrated tools, k

local-air-ollama
10 Jun 2026
Research

Exploration of Foundation Model-Based Robots in Patient and Elderly Care

DGX agent

arXiv:2606.10208v1 Announce Type: cross Abstract: Demand for older-adult and patient care is growing rapidly as populations age worldwide. Foundation models are increasingly being integrated into robo

researcharxiv-cs-ai
10 Jun 2026
Research

Interactions Between Crosscoder Features: A Compact Proofs Perspective

DGX agent

arXiv:2606.09940v1 Announce Type: cross Abstract: Dictionary learning methods like Sparse Autoencoders (SAEs) and crosscoders attempt to explain a model by decomposing its activations into independent

researcharxiv-cs-ai
10 Jun 2026
Safety

LLM-Aided Joint Secrecy Precoding and Trajectory for RSMA-Based Heterogeneous UAV Networks

DGX agent

arXiv:2507.17188v2 Announce Type: replace-cross Abstract: This paper investigates secure communications in rate-splitting multiple access (RSMA) enabled heterogeneous UAV networks, where multiple UAVs

safetyarxiv-cs-ai
10 Jun 2026
Tools

Markie is one of my favorite people. I have learned so much from her perspective that is equal parts soulful clarity about what computing sh…

DGX agent

Markie is one of my favorite people. I have learned so much from her perspective that is equal parts soulful clarity about what computing should aspire to be for humans, and the demanding complexity t

toolslinus-lee--x
10 Jun 2026
Research

Model-Based Reinforcement Learning in Discrete-Action Non-Markovian Reward Decision Processes

DGX agent

arXiv:2512.14617v2 Announce Type: replace-cross Abstract: Many practical decision-making problems involve tasks whose success depends on the entire system history, rather than on achieving a state wit

researcharxiv-cs-ai
10 Jun 2026
Model Releases

Monte Carlo Pass Search: Using Trajectory Generation for 3D Counterfactual Pass Evaluation in Football

DGX agent

arXiv:2606.11120v1 Announce Type: new Abstract: We recast pass evaluation in football (soccer) as a Monte Carlo Tree Search (MCTS)-like evaluation problem whose components mostly exist in the literatu

model-releasesarxiv-cs-ai
10 Jun 2026
Local Ai

RunPod AI Hub - Public Beta 1.34 live

DGX agent

RunPod Hub is a centralized catalog of preconfigured AI repositories that you can browse, deploy, and share, optimized for RunPod's Serverless infrastructure to deploy in minutes. The platform include

local-air-stablediffusion
10 Jun 2026
Model Releases

SCOPE: Sequential Causal Optimization of Process Interventions

DGX agent

arXiv:2512.17629v4 Announce Type: replace-cross Abstract: Prescriptive Process Monitoring (PresPM) recommends interventions during running business processes to optimize key performance indicators (KP

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

The Shibboleth Effect: Auditing the Cross-Lingual Distributional Skew of Large Language Models

DGX agent

arXiv:2606.11082v1 Announce Type: new Abstract: This study investigates cross-lingual distributional skew (the Shibboleth Effect) in frontier large language models (LLMs) subjected to sustained advers

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

This is a great article on how startups/frontier labs can coexist. Another way to look at this is task complexity - the number of bits of in…

DGX agent

This is a great article on how startups/frontier labs can coexist. Another way to look at this is task complexity - the number of bits of information needed to specify a task such that AI can solve th

model-releasesjerry-liu--x
10 Jun 2026
Safety

Using Probabilistic Programs to Train Inductive Reasoning in Large Language Models

DGX agent

arXiv:2606.09856v1 Announce Type: cross Abstract: Post-training Large Language Models (LLMs) for reasoning typically focuses on deductive tasks such as mathematics and coding where correctness is veri

safetyarxiv-cs-ai
10 Jun 2026
Model Releases

When RL Fails after SFT: Rejuvenating Model Plasticity for Robust SFT-to-RL Handoff

DGX agent

arXiv:2606.09932v1 Announce Type: cross Abstract: Supervised Fine-Tuning (SFT) followed by Reinforcement Learning (RL) has become a standard pipeline for Large Language Model (LLM) post-training. SFT

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

ACTIVE-o3: Empowering MLLMs with Active Perception via Pure Reinforcement Learning

DGX agent

arXiv:2505.21457v2 Announce Type: replace-cross Abstract: Active vision, also known as active perception, refers to actively selecting where and how to look in order to gather task-relevant informatio

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

An Agency-Transferring Model-Free Policy Enhancement Technique

DGX agent

arXiv:2606.09825v1 Announce Type: cross Abstract: Training reinforcement learning (RL) policies from scratch is costly: it requires careful reward and environment design, extensive tuning, and substan

safetyarxiv-cs-ai
9 Jun 2026
Local Ai

An Alternative Trajectory for Generative AI

DGX agent

arXiv:2603.14147v2 Announce Type: replace Abstract: The generative artificial intelligence (AI) ecosystem is undergoing rapid transformations that threaten its sustainability. As models transition fro

local-aiarxiv-cs-ai
9 Jun 2026
Model Releases

Artificial Intelligence for Mathematical Reasoning: An Integrated Survey of Language Models, Neuro-symbolic Systems, and Verified Discovery

DGX agent

arXiv:2606.08728v1 Announce Type: new Abstract: Mathematical reasoning has long served as a stringent test of machine intelligence; over the past decade, it has moved from a niche problem within NLP t

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

BREAKING: Anthropic just dropped Claude Fable 5—this is Mythos, made safe for public release. It is the best coding model in the world. We'v…

DGX agent

BREAKING: Anthropic just dropped Claude Fable 5—this is Mythos, made safe for public release. It is the best coding model in the world. We've been testing it internally @every for the last week or so

model-releasessimon-willison--x
9 Jun 2026
Model Releases

Claude Code-Driving Scenario Mining for the Argoverse 2 Challenge

DGX agent

arXiv:2606.09180v1 Announce Type: new Abstract: We present our submission to the CVPR 2026 Argoverse 2 Scenario Mining Challenge. Our system uses a four-stage pipeline: (1) autonomous code generation

model-releasesarxiv-cs-cv
9 Jun 2026
Safety

Code Is More Than Text: Uncertainty Estimation for Code Generation

DGX agent

arXiv:2606.09577v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as code generators, where silently wrong programs pose real safety and reliability risks. Relia

safetyarxiv-cs-lg
9 Jun 2026
Safety

DIVERGE: Diversity-Enhanced RAG for Open-Ended Information Seeking

DGX agent

arXiv:2602.00238v2 Announce Type: replace-cross Abstract: Existing retrieval-augmented generation (RAG) systems often assume that each query has a single correct answer. This assumption overlooks open

safetyarxiv-cs-ai
9 Jun 2026
Safety

Evaluating AI Investment Strategies

DGX agent

arXiv:2606.08791v1 Announce Type: cross Abstract: We study the problem of auditing a black-box algorithmic decision-maker from observable inputs and outputs alone. Our main result is an exact decompos

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

🚨 Fable 5 is something to pay attention to. This is another 'I had early access to the new Mythos-class Anthropic model and I want to tell …

DGX agent

🚨 Fable 5 is something to pay attention to. This is another 'I had early access to the new Mythos-class Anthropic model and I want to tell you what I thought of it' post. I know, I know, it's annoying

model-releasesallie-k--miller--x
9 Jun 2026
Model Releases

Fable 5 is the biggest step up I’ve felt in our models since Opus 4.5 back in November. After 4.5 came out I uninstalled my IDE when I reali…

DGX agent

Fable 5 is the biggest step up I’ve felt in our models since Opus 4.5 back in November. After 4.5 came out I uninstalled my IDE when I realized that I’d been doing 100% of my coding in a terminal for

model-releasesboris-cherny--x
9 Jun 2026
Local Ai

Federated Large Language Models: Current Progress and Future Directions

DGX agent

arXiv:2409.15723v3 Announce Type: replace Abstract: Large Language Models have achieved impressive performance across diverse applications, yet their training typically depends on centralized data col

local-aiarxiv-cs-lg
9 Jun 2026
Research

Formalizing Learning from Language Feedback with Provable Guarantees

DGX agent

arXiv:2506.10341v2 Announce Type: replace Abstract: Interactively learning from observation and language feedback is an increasingly studied area driven by the emergence of large language model (LLM)

researcharxiv-cs-lg
9 Jun 2026
Model Releases

HA-VLN 2.0: An Open Benchmark and Leaderboard for Human-Aware Navigation in Discrete and Continuous Environments with Dynamic Multi-Human Interactions

DGX agent

arXiv:2503.14229v4 Announce Type: replace Abstract: Vision-and-Language Navigation (VLN) has been studied mainly in either discrete or continuous spaces, with little attention to dynamic, crowded envi

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

Harness Engineering for Physical AI: Robot Middleware Is the Harness Layer

DGX agent

arXiv:2606.09416v1 Announce Type: cross Abstract: Robot middleware faces a new role in the era of Physical AI. Learned policies, planners, and vision-language-action (VLA) models now enter deployed ro

safetyarxiv-cs-ai
9 Jun 2026
Applications

I want to share some thoughts about the impact of AI on org structure. This is based on early conversations with some of the top CHROs in Fo…

DGX agent

I want to share some thoughts about the impact of AI on org structure. This is based on early conversations with some of the top CHROs in Fortune 500 companies as well as other C-Suite leaders and sta

applicationsallie-k--miller--x
9 Jun 2026
Applications

iMaC: Translating Actions into Motion and Contact Images for Embodied World Models

DGX agent

arXiv:2606.09813v1 Announce Type: cross Abstract: Embodied world models have emerged as a pivotal paradigm for visual robotic decision-making and interactive environment simulation. However, conventio

applicationsarxiv-cs-cv
9 Jun 2026
Model Releases

Introducing Gemma 4 12B: a unified, encoder-free multimodal model

DGX agent

Gemma 4 12B is a unified, encoder-free multimodal model designed to bring high-performance intelligence to laptops and released under an Apache 2.0 license. It eliminates separate encoders by projecti

model-releasesgoogle-deepmind
9 Jun 2026
Safety

IR-SIM: A Lightweight Skill-Native Simulator for Navigation, Learning, and Benchmarking

DGX agent

arXiv:2606.08729v1 Announce Type: cross Abstract: Simulation plays a key role in automated robotics research supported by large language models (LLMs). However, existing simulators often require custo

safetyarxiv-cs-lg
9 Jun 2026
Tutorials

It's confusing that Trump constantly warns about the security of our elections & foreign interference, yet his administration has taken seve…

DGX agent

It's confusing that Trump constantly warns about the security of our elections & foreign interference, yet his administration has taken several steps to dismantle or divert resources away from key saf

tutorialsyann-lecun--x
9 Jun 2026
Safety

LUNA-AD: Lightweight Uncertainty-Aware Language Model with Lifelong Learning for Autonomous Driving

DGX agent

arXiv:2606.08470v1 Announce Type: new Abstract: While large language models (LLMs) offer promising reasoning capabilities, their integration into safety-critical driving systems is hindered by limited

safetyarxiv-cs-ro
9 Jun 2026
Model Releases

PRISM: PRior-guided Imagination Sampling in world Models

DGX agent

arXiv:2606.07974v1 Announce Type: cross Abstract: A learned world model provides a powerful physical intuition for evaluating future states. But its effectiveness in continuous control also depends cr

model-releasesarxiv-cs-ai
9 Jun 2026
← Previous
1…341342343344345…370
Next →