AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,770 results
Tools

Warelay -> OpenClaw

DGX agent

In preparation for a lightning talk I'm giving at PyCon US this afternoon I decided to figure out how many names OpenClaw has actually had since that first commit back in November. Thanks to this firs

toolssimon-willison
16 May 2026
Applications

FrontierSmith: Synthesizing Open-Ended Coding Problems at Scale

DGX agent

arXiv:2605.14445v1 Announce Type: new Abstract: Many real-world coding challenges are open-ended and admit no known optimal solution. Yet, recent progress in LLM coding has focused on well-defined tas

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
applicationsarxiv-cs-lg
15 May 2026
Safety

Adaptive Smooth Tchebycheff Attention for Multi-Objective Policy Optimization

DGX agent

arXiv:2605.12771v1 Announce Type: cross Abstract: Multi-objective reinforcement learning in robotic domains requires balancing complex, non-convex trade-offs between conflicting objectives. While line

safetyarxiv-cs-ai
14 May 2026
Model Releases

Fireworks Training Platform continues to expand. Today GLM 5.1 LoRA RL is now live via Training API: SFT, DPO, and full RL on a 200K context…

DGX agent

Fireworks Training Platform continues to expand. Today GLM 5.1 LoRA RL is now live via Training API: SFT, DPO, and full RL on a 200K context window → custom loss functions or smart defaults. No usage

model-releasesfireworks-ai--x
14 May 2026
Safety

Flow Matching for Offline Reinforcement Learning with Discrete Actions

DGX agent

arXiv:2602.06138v2 Announce Type: replace Abstract: Generative policies based on diffusion models and flow matching have shown strong promise for offline reinforcement learning (RL), but their applica

safetyarxiv-cs-lg
14 May 2026
Safety

In-Situ Behavioral Evaluation for LLM Fairness, Not Standardized-Test Scores

DGX agent

arXiv:2605.12530v1 Announce Type: cross Abstract: LLM fairness should be evaluated through in-situ conversational behavior rather than standardized-test Q&A benchmarks. We show that the standardized-t

safetyarxiv-cs-ai
14 May 2026
Safety

interwhen: A Generalizable Framework for Steering Reasoning Models with Test-time Verification

DGX agent

arXiv:2602.11202v3 Announce Type: replace-cross Abstract: Reasoning models produce long traces of intermediate decisions and tool calls, making test-time verification important for ensuring correctnes

safetyarxiv-cs-ai
14 May 2026
Local Ai

Ring-2.6-1T Open sourced today! Soooo looking forward to trying it on Ollama!

DGX agent

Ring-2.6-1T is a trillion-parameter flagship reasoning model designed for real-world complex task scenarios, now available as an open-source model. The model features about 63B activated parameters pe

local-air-ollama
14 May 2026
Model Releases

hey surprise - you can just launch interactive in tmux and then tail the jsonl - shipped a small wrapper...ralph loop iterating to full pari…

DGX agent

hey surprise - you can just launch interactive in tmux and then tail the jsonl - shipped a small wrapper...ralph loop iterating to full parity rn https://github.com/dexhorthy/shannon Starting June 15,

model-releasesjeremy-howard--x
13 May 2026
Model Releases

If you use any of the following with your Claude sub, your usage must got cut by 25x: - T3 Code - Conductor - zed - jean - “Claude -p” in yo…

DGX agent

If you use any of the following with your Claude sub, your usage must got cut by 25x: - T3 Code - Conductor - zed - jean - “Claude -p” in your ci - scripts to call Claude code from other tools They’re

model-releasesjeremy-howard--x
13 May 2026
Safety

Leveraging RAG for Training-Free Alignment of LLMs

DGX agent

arXiv:2605.11217v1 Announce Type: new Abstract: Large language model (LLM) alignment algorithms typically consist of post-training over preference pairs. While such algorithms are widely used to enabl

safetyarxiv-cs-lg
13 May 2026
Model Releases

Our continued commitment to Chromebooks, and looking ahead

DGX agent

At the Android Show yesterday, we introduced Googlebooks: a new category of premium laptops built with Gemini’s helpfulness at the core. Designed for Gemini Intelligence, Googlebooks will give persona

model-releasesgoogle-cloud-ai
13 May 2026
Model Releases

AlignDrive: Aligned Lateral-Longitudinal Planning for End-to-End Autonomous Driving

DGX agent

arXiv:2601.01762v2 Announce Type: replace-cross Abstract: Practical autonomous driving requires models that generalize by reasoning through spatial-temporal possibilities to exclude unsafe outcomes. W

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

AlphaExploitem: Going Beyond the Nash Equilibrium in Poker by Learning to Exploit Suboptimal Play

DGX agent

arXiv:2605.09150v1 Announce Type: new Abstract: Poker is an imperfect information game that has served as a long-standing benchmark for decision-making under uncertainty. To maximize utility beyond th

model-releasesarxiv-cs-lg
12 May 2026
Safety

Balancing Efficiency and Fairness in Traffic Light Control through Deep Reinforcement Learning

DGX agent

arXiv:2605.10170v1 Announce Type: new Abstract: Urban traffic congestion presents a significant challenge for modern cities, which impacts mobility and sustainability. Traditional traffic light contro

safetyarxiv-cs-lg
12 May 2026
Research

Communicating Sound Through Natural Language

DGX agent

arXiv:2605.08750v1 Announce Type: cross Abstract: Natural language is widely used to describe, prompt, and control audio systems, but rarely serves as the representation carrying audio itself. We intr

researcharxiv-cs-ai
12 May 2026
Tutorials

EFGCL: Learning Dynamic Motion through Spotting-Inspired External Force Guided Curriculum Learning

DGX agent

arXiv:2605.10063v1 Announce Type: new Abstract: Learning dynamic whole-body motions for legged robots through reinforcement learning (RL) remains challenging due to the high risk of failure, which mak

tutorialsarxiv-cs-ro
12 May 2026
Safety

Governing AI-Assisted Security Operations: A Design Science Framework for Operational Decision Support

DGX agent

arXiv:2605.09534v1 Announce Type: cross Abstract: Engineering managers increasingly must decide how to introduce generative artificial intelligence (AI), retrieval-augmented generation, and coding age

safetyarxiv-cs-ai
12 May 2026
Safety

Intelligent Autonomous Orchestration for Distributed Cloud Resources using Complex-Stability Analysis

DGX agent

arXiv:2605.08139v1 Announce Type: cross Abstract: In modern distributed cloud environments, efficient resource allocation is required as traditional scaling mechanisms are often subject to cloud thras

safetyarxiv-cs-ai
12 May 2026
Model Releases

Is Your Driving World Model an All-Around Player?

DGX agent

arXiv:2605.10858v1 Announce Type: new Abstract: Today's driving world models can generate remarkably realistic dash-cam videos, yet no single model excels universally. Some generate photorealistic tex

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

LEAF-SQL: Level-wise Exploration with Adaptive Fine-graining for Text-to-SQL Skeleton Prediction

DGX agent

arXiv:2605.09295v1 Announce Type: new Abstract: Text-to-SQL translates natural language questions into executable SQL queries, enabling intuitive database access for non-experts. While large language

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Learning from Trials and Errors: Reflective Test-Time Planning for Embodied LLMs

DGX agent

arXiv:2602.21198v2 Announce Type: replace-cross Abstract: Embodied LLMs endow robots with high-level task reasoning, but they cannot reflect on what went wrong or why, turning deployment into a sequen

model-releasesarxiv-cs-ai
12 May 2026
Safety

NEXUS: Continual Learning of Symbolic Constraints for Safe and Robust Embodied Planning

DGX agent

arXiv:2605.09387v1 Announce Type: new Abstract: While Large Language Models (LLMs) have catalyzed progress in embodied intelligence, a fundamental gap between their inherent probabilistic uncertainty

safetyarxiv-cs-ai
12 May 2026
Model Releases

parameter golf was a blast. 2,000+ submissions. 1,000+ verified github accounts. ideas ranging from quantization and depth recurrence to TTT…

DGX agent

parameter golf was a blast. 2,000+ submissions. 1,000+ verified github accounts. ideas ranging from quantization and depth recurrence to TTT LoRA, SSMs, H-nets, JEPA, and more. autoresearch made itera

model-releasesopenai--x
12 May 2026
Model Releases

ProactBench: Beyond What The User Asked For

DGX agent

arXiv:2605.09228v1 Announce Type: cross Abstract: Most LLM benchmarks score how well a model responds to explicit requests. They leave unmeasured a different conversational ability: noticing and actin

model-releasesarxiv-cs-ai
12 May 2026
Safety

Reinforcement Learning for Scalable and Trustworthy Intelligent Systems

DGX agent

arXiv:2605.08378v1 Announce Type: cross Abstract: Reinforcement learning has become a powerful paradigm for improving the capability of intelligent systems, but its practical deployment faces two cent

safetyarxiv-cs-ai
12 May 2026
Local Ai

UniUncer: Unified Dynamic Static Uncertainty for End to End Driving

DGX agent

arXiv:2603.07686v2 Announce Type: replace-cross Abstract: End-to-end (E2E) driving has become a cornerstone of both industry deployment and academic research, offering a single learnable pipeline that

local-aiarxiv-cs-cv
12 May 2026
Research

From Standalone LLMs to Integrated Intelligence: A Survey of Compound Al Systems

DGX agent

arXiv:2506.04565v2 Announce Type: replace-cross Abstract: Compound AI Systems (CAIS) are an emerging paradigm that integrates large language models (LLMs) with external components, including retriever

researcharxiv-cs-cl
11 May 2026
Model Releases

GazeVLM: Active Vision via Internal Attention Control for Multimodal Reasoning

DGX agent

arXiv:2605.07817v1 Announce Type: cross Abstract: Human visual reasoning is governed by active vision, a process where metacognitive control drives top-down goal-directed attention, dynamically routin

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Graph Representation Learning Augmented Model Manipulation on Federated Fine-Tuning of LLMs

DGX agent

arXiv:2605.07961v1 Announce Type: new Abstract: Federated fine-tuning (FFT) has emerged as a privacy-preserving paradigm for collaboratively adapting large language models (LLMs). Built upon federated

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

MathlibPR: Pull Request Merge-Readiness Benchmark for Formal Mathematical Libraries

DGX agent

arXiv:2605.07147v1 Announce Type: cross Abstract: The ecosystem of Lean and Mathlib has become the de facto standard for large language model (LLM) assisted formal reasoning with remarkable successes

model-releasesarxiv-cs-ai
11 May 2026
Safety

Reason to Play: Behavioral and Brain Alignment Between Frontier LRMs and Human Game Learners

DGX agent

arXiv:2605.08019v1 Announce Type: new Abstract: Humans rapidly learn abstract knowledge when encountering novel environments and flexibly deploy this knowledge to guide efficient and intelligent actio

safetyarxiv-cs-ai
11 May 2026
Research

Scalable Option Learning in High-Throughput Environments

DGX agent

arXiv:2509.00338v3 Announce Type: replace-cross Abstract: Hierarchical reinforcement learning (RL) has the potential to enable effective decision-making over long timescales. Existing approaches, whil

researcharxiv-cs-ai
11 May 2026
Model Releases

SCENE: Recognizing Social Norms and Sanctioning in Group Chats

DGX agent

arXiv:2605.07823v1 Announce Type: new Abstract: Online group chats are social spaces with implicit behavior patterns that, when broken, are often met with social sanctioning from the group. The abilit

model-releasesarxiv-cs-cl
11 May 2026
Local Ai

Self Driving Datasets: From 20 Million Papers to Nuanced Biomedical Knowledge at Scale

DGX agent

arXiv:2605.07022v1 Announce Type: new Abstract: Manually curated biomedical repositories -- spanning bioactivity, genomics, and chemistry -- are expensive to maintain, lag behind primary literature, a

local-aiarxiv-cs-lg
11 May 2026
Research

Towards Highly-Constrained Human Motion Generation with Retrieval-Guided Diffusion Noise Optimization

DGX agent

arXiv:2605.08054v1 Announce Type: new Abstract: Generating human motion that satisfies customized zero-shot goal functions, enabling applications such as controllable character animation and behavior

researcharxiv-cs-cv
11 May 2026
Research

Deco: Extending Personal Physical Objects into Pervasive AI Companion through a Dual-Embodiment Framework

DGX agent

arXiv:2605.03882v1 Announce Type: cross Abstract: Individuals frequently form deep attachments to physical objects (e.g., plush toys) that usually cannot sense or respond to their emotions. While AI c

researcharxiv-cs-ai
7 May 2026
Model Releases

Information Coordination as a Bridge: A Neuro-Symbolic Architecture for Reliable Autonomous Driving Scene Understanding

DGX agent

arXiv:2605.04475v1 Announce Type: new Abstract: Reliable autonomous driving requires scene understanding that is semantically consistent across heterogeneous sensors and verifiable at the reasoning st

model-releasesarxiv-cs-cv
7 May 2026
Model Releases

LUCAS-MEGA: A Large-Scale Multimodal Dataset for Representation Learning in Soil-Environment Systems

DGX agent

arXiv:2605.04323v1 Announce Type: new Abstract: Understanding soil is fundamental to agriculture, carbon cycling, and environmental sustainability, yet progress is limited by fragmented and heterogene

model-releasesarxiv-cs-lg
7 May 2026
Local Ai

MongoDB announces platform enhancements for enterprise-ready AI production

DGX agent

Popular NoSQL-based database company MongoDB Inc. today announced a new set of capabilities during the company’s .Local conference in London, bringing together everything software and artificial intel

local-aisiliconangle
7 May 2026
Safety

Safety Must Precede the Deployment of Open-Ended AI

DGX agent

arXiv:2502.04512v3 Announce Type: replace Abstract: AI advancements have been significantly driven by a combination of foundation models and curiosity-driven learning aimed at increasing capability an

safetyarxiv-cs-ai
7 May 2026
Model Releases

Saw this and thought 'yes! ChatGPT voice mode is going to stop acting like a two-year-model' but that upgrade hasn't shipped just yet

DGX agent

Saw this and thought 'yes! ChatGPT voice mode is going to stop acting like a two-year-model' but that upgrade hasn't shipped just yet Introducing GPT-Realtime-2 in the API: our most intelligent voice

model-releasessimon-willison--x
7 May 2026
Hardware

A great conversation between Noah Kravitz from the @nvidia team + @hwchase17.

DGX agent

A great conversation between Noah Kravitz from the @nvidia team + @hwchase17. “Every enterprise needs a claw strategy.” How did @LangChain go from a weekend project to 1B+ downloads in 3 years? We sat

hardwareharrison-chase--x
6 May 2026
Safety

Learning Reactive Dexterous Grasping via Hierarchical Task-Space RL Planning and Joint-Space QP Control

DGX agent

arXiv:2605.03363v1 Announce Type: new Abstract: In this work, we propose a hybrid hierarchical control framework for reactive dexterous grasping that explicitly decouples high-level spatial intent fro

safetyarxiv-cs-ro
6 May 2026
Model Releases

MCP-Atlas: A Large-Scale Benchmark for Tool-Use Competency with Real MCP Servers

DGX agent

arXiv:2602.00933v2 Announce Type: replace-cross Abstract: The Model Context Protocol (MCP) is rapidly becoming the standard interface for Large Language Models (LLMs) to discover and invoke external t

model-releasesarxiv-cs-ai
6 May 2026
Safety

MINT: Minimal Information Neuro-Symbolic Tree for Objective-Driven Knowledge-Gap Reasoning and Active Elicitation

DGX agent

arXiv:2602.05048v2 Announce Type: replace Abstract: Joint planning through language-based interactions is a key area of human-AI teaming. Planning problems in the open world often involve various aspe

safetyarxiv-cs-ai
6 May 2026
Model Releases

Optimal control of the future via prospective learning with control

DGX agent

arXiv:2511.08717v4 Announce Type: replace-cross Abstract: Optimal control of the future is the next frontier for AI. Current approaches to this problem are typically rooted in reinforcement learning (

model-releasesarxiv-cs-lg
6 May 2026
Tutorials

Preemptive Solving of Future Problems: Multitask Preplay in Humans and Machines

DGX agent

arXiv:2507.05561v2 Announce Type: replace Abstract: Humans can pursue a near-infinite variety of tasks, but typically can only pursue a small number at the same time. We hypothesize that humans levera

tutorialsarxiv-cs-lg
6 May 2026
← Previous
1…311312313314315…371
Next →