AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,972 results
Model Releases

Behavioral Inference at Scale: The Fundamental Asymmetry Between Motivations and Belief Systems

DGX agent

arXiv:2509.05624v3 Announce Type: replace-cross Abstract: How much information about an agent's underlying values can be recovered from its observable behavior? This question matters for any approach

model-releasesarxiv-cs-lg
12 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

Continuous Interaction Diffusion: A Diffusion-Native Runtime for Asynchronous Tool-Augmented Reasoning

DGX agent

arXiv:2608.10438v1 Announce Type: new Abstract: Large language models increasingly rely on external tools to access up-to-date information, perform computation, and interact with the outside world. Fo

agentsarxiv-cs-ai
12 Aug 2026
Hardware

Hand-Written PTX Tensor-Core GEMM Kernels: A Multi-Precision Study on NVIDIA L4

DGX agent

arXiv:2608.10103v1 Announce Type: cross Abstract: High-performance Tensor Core kernels rely on a low-level PTX pipeline built from asynchronous data movement with cp.async, warp-level matrix loads wit

hardwarearxiv-cs-ai
12 Aug 2026
Tutorials

Live now: our Memory & Continual Learning Track from AI Engineer World's Fair 2026. Thesis: we scaled intelligence and got the world's smart…

DGX agent

Live now: our Memory & Continual Learning Track from AI Engineer World's Fair 2026. Thesis: we scaled intelligence and got the world's smartest novice. https://www.youtube.com/watch?v=iqloyWCGYQQ&list

tutorialsswyx--x
12 Aug 2026
Agents

Nebius shares jump 34% on continued AI infrastructure demand

DGX agent

Shares of Nebius Group NV closed 34% higher today after it reported second-quarter earnings that topped expectations across the board. The Netherlands-based company operates a cloud platform optimized

agentssiliconangle
12 Aug 2026
Agents

The Deliberative Deficit: An Empirical Critique of LLMs in Democratic Discourse

DGX agent

arXiv:2608.10186v1 Announce Type: cross Abstract: LLMs are increasingly deployed in settings that require collective reasoning on complex, value-laden problems. Confidence in these deployments rests l

agentsarxiv-cs-ai
12 Aug 2026
Agents

A Communication-Efficient Digital Twin Framework for PSO-Based Swarm Navigation and Obstacle Avoidance

DGX agent

arXiv:2406.19930v4 Announce Type: replace Abstract: Swarm-based target localization in industrial environments faces two major challenges: navigating obstacle-rich spaces and managing intensive commun

agentsarxiv-cs-ro
11 Aug 2026
Model Releases

Benchmarking In-context Experiential Learning Through Repeated Product Recommendations

DGX agent

arXiv:2511.22130v2 Announce Type: replace Abstract: To navigate ever-shifting real-world environments, agents must grapple with incomplete knowledge and adapt their strategies through experience. Howe

model-releasesarxiv-cs-lg
11 Aug 2026
Agents

Had so much fun giving this talk at @sequoia about @harvey’s moneyball approach to building a research lab. The biggest mistake I made in th…

DGX agent

Had so much fun giving this talk at @sequoia about @harvey’s moneyball approach to building a research lab. The biggest mistake I made in the early days of Harvey was trying to play the Yankees baseba

agentssonya-huang--x
11 Aug 2026
Agents

IntelliAudit: Using Large Language Models to Evaluate Audit Controls

DGX agent

arXiv:2608.07688v1 Announce Type: new Abstract: IT audits require auditors to judge whether heterogeneous organizational evidence satisfies semantic security and compliance controls. This judgment is

agentsarxiv-cs-ai
11 Aug 2026
Agents

Skills in Weights, Memory in Code: Hybrid Learning for Memory-Dependent Robot Manipulation

DGX agent

arXiv:2608.09410v1 Announce Type: new Abstract: Modern vision-language-action (VLA) policies have acquired broad manipulation skills, but typically generate each action chunk from the current observat

agentsarxiv-cs-ro
11 Aug 2026
Agents

The Capability Ladder: A Curriculum-Modernization Framework for Workforce Readiness in the AI Era

DGX agent

arXiv:2608.07779v1 Announce Type: new Abstract: Artificial intelligence is changing the task composition of computing work faster than curricula and training typically adapt. This is a curriculum-fram

agentsarxiv-cs-ai
11 Aug 2026
Model Releases

TRACE: TRajectory Attribution for Automated Context Engineering

DGX agent

arXiv:2608.09153v1 Announce Type: new Abstract: Production AI agents fail when their context sources -- system prompts, knowledge bases, tool descriptions, and procedural skills -- contain errors or g

model-releasesarxiv-cs-ai
11 Aug 2026
Agents

VDGR-RAG: Vectors, Directories, Graphs, and Reflection Are All You Need for Unified Reasoning over Hierarchical Enterprise Knowledge

DGX agent

arXiv:2608.07994v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) is essential for enterprise knowledge question answering (QA), particularly in domains with complex product documen

agentsarxiv-cs-ai
11 Aug 2026
Research

Interaction Creates Dynamical AI Behavior Absent in Isolation

DGX agent

arXiv:2608.07457v1 Announce Type: new Abstract: What will happen when AI agents interact in daily life, e.g. when one AI starts bossing another around? We find a counterintuitive answer that opens new

researcharxiv-cs-ai
10 Aug 2026
Agents

Probing Visual Concepts in Lightweight Vision-Language Models for Automated Driving

DGX agent

arXiv:2603.06054v2 Announce Type: replace-cross Abstract: The use of Vision-Language Models (VLMs) in automated driving applications is becoming increasingly common, with the aim of leveraging their r

agentsarxiv-cs-ai
10 Aug 2026
Agents

Vehicle routing problem using deep reinforcement learning - A case study about truck planning in the industry

DGX agent

arXiv:2608.06668v1 Announce Type: new Abstract: As an important component of the supply chain industry, transportation has experienced rapid development in the past decade with the assistance of digit

agentsarxiv-cs-ai
10 Aug 2026
Model Releases

Auto mode is now the default in Claude Code for Pro, Max, and Team plans

DGX agent

Auto mode is now the default in Claude Code for Pro, Max, and Team plans Anthropic are really confident in Claude Code's auto mode, to the point that they are making it the default setting for new ses

model-releasessimon-willison
8 Aug 2026
Agents

Computationally Efficient Collaborative Communication Via Regularity-Based Coarsening

DGX agent

arXiv:2608.05327v1 Announce Type: cross Abstract: Our results show that the existence of a short high-utility protocol already suffices for efficient communication. In particular, in a game with n pos

agentsarxiv-cs-lg
7 Aug 2026
Safety

How Cohere Health digitizes clinical policies using Amazon Bedrock AgentCore

DGX agent

In this post, you learn how Cohere Health built a multi-tenant agentic architecture on AgentCore using AgentCore Runtime’s secure MicroVM isolation, unified tool access through AgentCore Gateway, Agen

safetyaws-ml-blog
7 Aug 2026
Model Releases

Recursive Synthesis for Long-Horizon Terminal Tasks

DGX agent

arXiv:2608.05466v1 Announce Type: new Abstract: High-quality long-horizon training data for terminal agents is expensive to produce, often costing hundreds to thousands of dollars per task, because ea

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

upgraded my stack, and i can now work on almost anything from anywhere hands free: - talk to chief of staff (via remote codex voice or text)…

DGX agent

upgraded my stack, and i can now work on almost anything from anywhere hands free: - talk to chief of staff (via remote codex voice or text) - chief assigns tasks to managers of various projects - man

model-releasesyohei-nakajima--x
7 Aug 2026
Agents

Zero-code, low-cost data ingestion: New BigQuery DTS capabilities

DGX agent

In a fast-paced digital economy, data is your most critical engine. Yet, many enterprises find themselves trapped in a costly paradox, spending over 100 hours a week building and fixing fragile, in-ho

agentsgoogle-cloud-ai
7 Aug 2026
Agents

i guess this is a good time to mention that smol forge is open for the first 100 alpha users. get your usernames! (tire kickers who dont mak…

DGX agent

i guess this is a good time to mention that smol forge is open for the first 100 alpha users. get your usernames! (tire kickers who dont make any commits will be kicked out by eod) point clanker to fo

agentsswyx--x
6 Aug 2026
Agents

Improving Auto-Design of Neural PDE Solvers with a Domain-Specific Language

DGX agent

arXiv:2608.04384v1 Announce Type: new Abstract: Neural PDE solver auto-design is fundamentally a search-space representation problem. In the space of unrestricted Python programs, valid solvers form a

agentsarxiv-cs-ai
6 Aug 2026
Agents

OctoLong: Mid-Training On Cross-Repository Code Contexts Enhances Long-Context Modeling

DGX agent

arXiv:2608.05141v1 Announce Type: new Abstract: Context lengths of language models (LMs) have dramatically increased, driven by the demands for in-context learning, self-improvement, and long-horizon

agentsarxiv-cs-ai
6 Aug 2026
Agents

When Shared Rollouts Fail in Defensive Driving Evaluation: A NAVSIM Score Basis Audit

DGX agent

arXiv:2608.04896v1 Announce Type: new Abstract: Defensive driving scores are useful only when they preserve distinctions between policies that observe surrounding actors and those that do not. Re-simu

agentsarxiv-cs-ai
6 Aug 2026
Research

EFX Allocation In (Multi)Hypergraphs

DGX agent

arXiv:2608.03171v1 Announce Type: cross Abstract: We study fair allocations of indivisible goods among agents with heterogeneous monotone valuations. As fair we consider the allocations that are envy-

researcharxiv-cs-ai
5 Aug 2026
Safety

From Routes to Steps: Separating Semantic Progress from Local Execution in Vision-and-Language Navigation

DGX agent

arXiv:2608.03143v1 Announce Type: new Abstract: Vision-and-Language Navigation (VLN) requires an agent to follow a route-level instruction by executing its constituent steps from egocentric visual obs

safetyarxiv-cs-cv
5 Aug 2026
Model Releases

Getting the Parameters Right: A Difficulty-Graded Benchmark and Probe-Guided Training for LLM Tool Calls

DGX agent

arXiv:2608.03071v1 Announce Type: new Abstract: Large language model agents derive much of their capability from tool use. Existing research on tool use has largely focused on selecting the right tool

model-releasesarxiv-cs-ai
5 Aug 2026
Agents

'OCR is just a feature now. Frontier models will eat it.' We hear this constantly. The data says otherwise. Across three GPT generations, pa…

DGX agent

'OCR is just a feature now. Frontier models will eat it.' We hear this constantly. The data says otherwise. Across three GPT generations, parsing accuracy gained ~24 points, while cost per page 4x'd.

agentsjerry-liu--x
5 Aug 2026
Agents

PAIChecker: Uncovering and Checking PR-Issue Misalignment in SWE-Bench-Like Benchmarks

DGX agent

arXiv:2607.28587v2 Announce Type: replace-cross Abstract: SWE-bench-like benchmarks are widely used for evaluating LLM's issue resolution capability. They typically follow a common construction pipeli

agentsarxiv-cs-ai
5 Aug 2026
Agents

Residual Flow Matching with Dynamic Cross-Interaction for 3D Multi-Person Motion Prediction

DGX agent

arXiv:2608.03379v1 Announce Type: new Abstract: 3D multi-person motion prediction requires modeling both individual kinematics and inter-person interactions. While Flow Matching is effective for multi

agentsarxiv-cs-cv
5 Aug 2026
Safety

AI-Based Thesis Assessment: An Empirical Study of Human Evaluation Priorities and Their Impact on Automated Assessment

DGX agent

arXiv:2608.00717v1 Announce Type: cross Abstract: Rubric-based AI systems for thesis assessment use criterion weights to assign different levels of importance to evaluation criteria. These weights are

safetyarxiv-cs-cl
4 Aug 2026
Applications

Don't Offer What Can't Be Done: Deterministic Executability Gating for LLM Skill Selection at Scale

DGX agent

arXiv:2608.01050v1 Announce Type: cross Abstract: Production LLM agents that select from large skill libraries face a limitation that semantic relevance alone cannot resolve: a skill may match a user'

applicationsarxiv-cs-cl
4 Aug 2026
Model Releases

Humans Are More Diverse: Frontier LLMs Show Extreme Policies in Idealised AI Development Races

DGX agent

arXiv:2608.01193v1 Announce Type: cross Abstract: An AI development race creates a multi-agent safety dilemma. Each company can develop slowly and safely, or move faster while taking a risk that may r

model-releasesarxiv-cs-lg
4 Aug 2026
Safety

Inference-Time Policy Alignment for Fair Reinforcement Learning

DGX agent

arXiv:2608.00175v1 Announce Type: new Abstract: Deep reinforcement learning (RL) agents achieve strong performance by optimizing scalar reward functions. However, once deployed, the policies of these

safetyarxiv-cs-lg
4 Aug 2026
Agents

PB^2: Preference Space Exploration via Population-Based Methods in Preference-Based Reinforcement Learning

DGX agent

arXiv:2506.13741v2 Announce Type: replace-cross Abstract: Preference-based reinforcement learning (PbRL) has emerged as a promising approach for learning behaviors from human feedback without predefin

agentsarxiv-cs-lg
4 Aug 2026
Model Releases

PredAct-Bench: Benchmarking Tool-Augmented Dialogue under Controlled Tool Noise

DGX agent

arXiv:2608.02372v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed in task-oriented dialogue systems that support multi-step decision-making in high-stakes domains

model-releasesarxiv-cs-cl
4 Aug 2026
Agents

RubricReviewer: From Direct Critique to Objective and Comprehensive Rubric-Driven Peer Review

DGX agent

arXiv:2608.00005v1 Announce Type: new Abstract: Peer review at major venues is under unprecedented submission pressure, motivating the use of large language models (LLMs) as review assistants. Existin

agentsarxiv-cs-cl
4 Aug 2026
Agents

We were proud to host “Build with Frontier Intelligence,” http://Z.ai ’s first community meetup, together with @AISingapore, at Tencent’s ve…

DGX agent

We were proud to host “Build with Frontier Intelligence,” http://Z.ai ’s first community meetup, together with @AISingapore, at Tencent’s venue. During the event, http://Z.ai shared insights into the

agentszhipu-ai--x
4 Aug 2026
Model Releases

Data Turnstile: A Scalable Open Framework for Function-Calling Data Generation

DGX agent

arXiv:2607.29250v1 Announce Type: new Abstract: Small language models (SLMs) are attractive for agentic deployment due to low latency, reduced cost, and on-device privacy, yet they struggle with tool-

model-releasesarxiv-cs-cl
3 Aug 2026
Model Releases

DungeonBench: A Benchmark for Rules-Rich Tactical Reasoning in Dungeons & Dragons Combat

DGX agent

arXiv:2607.29577v1 Announce Type: new Abstract: Games and simulators make valuable benchmarks by turning decisions into measurable outcomes, but many current suites under-test rules-rich tactical reas

model-releasesarxiv-cs-ai
3 Aug 2026
Agents

UltraSAM3: A Concept-Driven Foundation Model for Universal Ultrasound Image Segmentation

DGX agent

arXiv:2607.29200v1 Announce Type: new Abstract: Ultrasound imaging has become increasingly widespread in clinical practice due to its portability, low cost and real-time capability, making ultrasound

agentsarxiv-cs-cv
3 Aug 2026
Safety

ViSAGE: Constructing Self-Correcting Memories for Long-Form Video Understanding

DGX agent

arXiv:2607.28678v1 Announce Type: new Abstract: Multimodal agents operating in long-horizon environments must build and continually update multimedia memories to support entity-consistent, temporally

safetyarxiv-cs-ai
3 Aug 2026
Agents

OpenAI really ought give @HuggingFace a $100M grant, much as @ClementDelangue is asking. And anyone who wants to understand the attack from …

DGX agent

OpenAI really ought give @HuggingFace a $100M grant, much as @ClementDelangue is asking. And anyone who wants to understand the attack from HF’s side should read this. Great walkthrough; A+ for visual

agentsgary-marcus--x
1 Aug 2026
Model Releases

Albilich: Steerable Proof-State Orchestration for LLM-Based Mathematical Research with CAS Integration

DGX agent

arXiv:2607.27705v1 Announce Type: cross Abstract: Large language models can contribute useful ideas to mathematical research, yet long-horizon proof attempts remain difficult to coordinate, evaluate,

model-releasesarxiv-cs-lg
31 Jul 2026
Local Ai

AlphaSchema: Exploring the Space of Trading Semantics for LLM-Based Alpha Mining

DGX agent

arXiv:2607.26642v1 Announce Type: new Abstract: Automated alpha mining has increasingly adopted large language model (LLM) agents for factor generation and iterative discovery. However, existing LLM-b

local-aiarxiv-cs-ai
31 Jul 2026
← Previous
1…217218219220221…375
Next →