AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,959 results
9 Jul 2026

Built from scratch by Grok 4.5 + Grok Build in UE5.8: a cyberpunk L-corner street with neon facades, rain, signs, and crowds walking through…

AgentsDGX agent

Built from scratch by Grok 4.5 + Grok Build in UE5.8: a cyberpunk L-corner street with neon facades, rain, signs, and crowds walking through the scene. End-to-end, the run took 10.75M tokens, 36.5 min

CoMind: Understanding Collaborative Human Activity from Multiple Minds and Views

AgentsDGX agent

arXiv:2607.06691v1 Announce Type: new Abstract: Human-human collaboration is a fundamental aspect of everyday life, essential to success in a wide range of goal-directed activities from household task

HiFuzz: Hierarchical Reinforcement Learning for Semantic-Aware and Adaptive CPU Fuzzing

AgentsDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.06619v1 Announce Type: cross Abstract: Modern processor verification struggles to reach deep architectural states due to the inefficiencies of traditional mutation-based fuzzing. We propose

new meta model

Model ReleasesDGX agent

new meta model (2) Muse Spark 1.1 is strongest at agentic performance, tool use, and computer use. It does well on long-running tasks with 1M token context window, can delegate execution to sub-agents

Power and Limitations of Aggregation in Compound AI Systems

AgentsDGX agent

arXiv:2602.21556v2 Announce Type: replace Abstract: When designing compound AI systems, a common approach is to query multiple copies of the same model and aggregate the responses to produce a synthes

Shared Modular Recurrence in Contextual MDPs for Universal Morphology Control

AgentsDGX agent

arXiv:2506.08630v3 Announce Type: replace Abstract: A universal controller for any robot morphology would greatly improve computational and data efficiency. Steps have been made towards such multi-rob

this is going to be a fun one! clay is one of the most ai-forward thinking companies i know of lots to learn here if you're in NYC

AgentsDGX agent

this is going to be a fun one! clay is one of the most ai-forward thinking companies i know of lots to learn here if you're in NYC Join us for a LangChain + Clay meetup with @palashshah, @jeffbarg, Vy

8 Jul 2026

Highly-recommended read. Aligns with what I see in my own harness: > Pi harness got the same success rate as harnesses from the LLM vendors …

AgentsDGX agent

Highly-recommended read. Aligns with what I see in my own harness: > Pi harness got the same success rate as harnesses from the LLM vendors with Opus and GPT, but at 2x less cost > GLM 5.2 was a major

Tangent classes of matroids and wonderful compactifications

AgentsDGX agent

arXiv:2607.05835v1 Announce Type: cross Abstract: For every loopless matroid M and every Feichtner--Yuzvinsky building set G containing the top flat, we construct an integral tangent class T_{M,G}^{Z}

7 Jul 2026

AI Innovators Adopt NVIDIA Vera — Why Max Single-Threaded CPU at Scale Matters

HardwareDGX agent

Max single-threaded CPUs at scale are a new category of CPUs built for the agentic AI era. Across the creation and deployment of an agentic system, the CPU is on the critical path for reasoning, respo

Auto: The AGI Compiler

Model ReleasesDGX agent

arXiv:2607.04542v1 Announce Type: cross Abstract: Every LLM agent run re-derives its behavior token by token on a frontier model: brilliant, expensive, slow, and unbounded. We present Auto, a compiler

Automated Data Readiness for Scientific AI

AgentsDGX agent

arXiv:2607.02771v1 Announce Type: new Abstract: Leadership computing facilities steward large-scale scientific datasets that routinely require substantial transformation before serving as AI training

Drive proactive security, prioritize risks with Google Threat Intelligence and Wiz ASM

AgentsDGX agent

Being more proactive continues to be a leading goal for security organizations. As AI accelerates the pace of vulnerability discovery and exploitation, organizations will rely on the personalization o

HunyuanOCR-1.5: Making Lightweight OCR VLMs Faster and Better

AgentsDGX agent

arXiv:2607.04884v1 Announce Type: new Abstract: We present HunyuanOCR-1.5, a lightweight end-to-end OCR-specialized vision-language model. HunyuanOCR unifies document parsing, text spotting, informati

LLM-Guided Transportation Hub Capacity Planning with Textual Business Inputs

AgentsDGX agent

arXiv:2607.03651v1 Announce Type: new Abstract: While traditional hub capacity planning models optimize effectively for quantitative inputs, they often fail to digest qualitative business context. We

Mask-based Predictive Representations for Reinforcement Learning

AgentsDGX agent

arXiv:2607.04153v1 Announce Type: cross Abstract: Vision-based deep reinforcement learning involves dealing with high-dimensional inputs of image information. It is crucial to abstract effective state

Policy Improvement with Style-Specific Demonstrations

SafetyDGX agent

arXiv:2506.16995v4 Announce Type: replace Abstract: Proficient game agents with diverse play styles enrich the gaming experience and enhance the replay value of games. However, recent advancements in

We've rolled out improvements to LlamaParse Cost Optimizer. Our intelligent tier routing now more reliably ensures you always strike the rig…

AgentsDGX agent

We've rolled out improvements to LlamaParse Cost Optimizer. Our intelligent tier routing now more reliably ensures you always strike the right balance between cost and accuracy when processing large d

What’s New in Microsoft Foundry | June 2026

Model ReleasesDGX agent

Claude is now generally available in Microsoft Foundry. Here's everything else that shipped between Build 2026 and the end of June — autopilot agents, expanded Toolboxes and Routines, Agent Optimizer'

4 Jul 2026

Come and chat with us in ICML 🥳 Excited to present Temporal Straightening for Latent Planning in the Tuesday morning session #1509. Let’s t…

AgentsDGX agent

Come and chat with us in ICML 🥳 Excited to present Temporal Straightening for Latent Planning in the Tuesday morning session #1509. Let’s talk about world models, JEPA and representation learning. Age

3 Jul 2026

Don't train the model, evolve the harness. I read a brilliant blog post from Hugging Face where they took a frozen open model scoring 0% on …

Model ReleasesDGX agent

Don't train the model, evolve the harness. I read a brilliant blog post from Hugging Face where they took a frozen open model scoring 0% on a hard legal agent benchmark, left its weights alone, and le

EvoPolicyGym: Evaluating Autonomous Policy Evolution in Interactive Environments

Model ReleasesDGX agent

arXiv:2607.02440v1 Announce Type: new Abstract: Autonomous agents are increasingly expected to improve executable policies through feedback, yet existing evaluations often collapse this process into a

FaithMed: Training LLMs For Faithful Evidence-Based Medical Reasoning

AgentsDGX agent

arXiv:2607.01440v1 Announce Type: new Abstract: Faithful reasoning is essential in medicine, where clinical decisions require transparent justification grounded in reliable evidence. Current medical L

Had an amazing time talking to some of the most energetic builders in AI at @aiDotEngineer thanks to @swyx for organizing and thanks to @alt…

AgentsDGX agent

Had an amazing time talking to some of the most energetic builders in AI at @aiDotEngineer thanks to @swyx for organizing and thanks to @altryne and friends for hallway chats and podcasts to share abo

Mastermind: Strategy-grounded Learning for Repository-Scale Vulnerability Reproduction

Model ReleasesDGX agent

arXiv:2607.01764v1 Announce Type: new Abstract: Repository-level vulnerability reproduction is a demanding software engineering (SE) task: an agent must inspect a codebase, infer the input grammar tha

PPTArena: A Benchmark for PowerPoint Editing

Model ReleasesDGX agent

arXiv:2512.03042v3 Announce Type: replace-cross Abstract: We introduce PPTArena, a benchmark for PowerPoint editing that evaluates how agents modify real slides from natural-language instructions. Unl

2 Jul 2026

Beyond Line of Sight: Hybrid Validation of V2X Collective Perception in Complex Scenarios

AgentsDGX agent

arXiv:2607.00874v1 Announce Type: new Abstract: This paper introduces a probabilistic framework and hybrid validation methodology for V2X-enabled Collective Perception (CP) in complex traffic scenario

Human-Machine Collaboration on Generative Meta-Learning: Model and Algorithm

AgentsDGX agent

arXiv:2607.00926v1 Announce Type: cross Abstract: Generalizing machine learning models to environments that differ from their training distribution remains a critical hurdle, particularly when data fr

1 Jul 2026

harvey is ripping!! honor to work with them

AgentsDGX agent

harvey is ripping!! honor to work with them Q2 recap for @harvey - +$100M NNARR - 53% DAU/MAU Key hires (including Q1) - Anique (CPO) - prev VP of Product at Rippling - Rachel (CMO) - prev CMO at Noti

NEW paper worth reading. (bookmark it) Autonomous research systems usually prove themselves on cherry-picked wins, human-framed topics, or a…

AgentsDGX agent

NEW paper worth reading. (bookmark it) Autonomous research systems usually prove themselves on cherry-picked wins, human-framed topics, or a handful of preset tasks. FARS runs the full loop at scale i

this video goes into more detail on how our integrations with @harborframework actually work! if the blog yesterday caught your eye definite…

AgentsDGX agent

this video goes into more detail on how our integrations with @harborframework actually work! if the blog yesterday caught your eye definitely watch this! thanks for all the help @kobe0938 @alexgshaw!

Xiaomi-GUI-0 Technical Report

Model ReleasesDGX agent

arXiv:2606.31410v1 Announce Type: new Abstract: Graphical user interface (GUI) agents build on vision-language models to complete user tasks end-to-end in real applications through interface actions s

30 Jun 2026

Exit-and-Join Dynamics and Equilibrium in Continuum Cooperative Games

AgentsDGX agent

arXiv:2606.28824v1 Announce Type: cross Abstract: This paper develops a continuum theory of exit-and-join coalition dynamics in nonatomic cooperative games. We extend the Aumann-Shapley value and the

good description of how someone has done wiki memory in the wild

AgentsDGX agent

good description of how someone has done wiki memory in the wild @hwchase17 @cognition @FactoryAI @karpathy We implemented wiki idea by hosting the 'files' as doc db objects, added an API for db query

Had a lot of fun at @aiDotEngineer talking about BM25!

AgentsDGX agent

Had a lot of fun at @aiDotEngineer talking about BM25! Cool Search paradigm by @jobergum: “BM25 + Grep is all you need”: 1. Use BM25 to narrow down a large universe of documents to a few candidates 2.

More time to build with Step 3.7 Flash: in partnership with @StepFun_ai, we’re extending the free usage period in Nous Portal by an addition…

AgentsDGX agent

More time to build with Step 3.7 Flash: in partnership with @StepFun_ai, we’re extending the free usage period in Nous Portal by an additional 15 days! Media Step 3.7 Flash is now free for 30 days via

Persona-Trained Monte Carlo: Estimating Market-Outcome Distributions via Swarms of Persona-Conditioned Neural Policy Bots in a Limit Order Book

SafetyDGX agent

arXiv:2606.29556v1 Announce Type: new Abstract: We propose Persona-Trained Monte Carlo (PTMC), a method for estimating distributions of market-outcome statistics by repeatedly simulating limit-order-b

Reinforcement Learning for Software Vulnerability Analysis: A Systematic Review with Emphasis on C/C++ Source Code and Static Analysis

AgentsDGX agent

arXiv:2606.28403v1 Announce Type: cross Abstract: Vulnerability detection in C/C++ software remains a major security challenge due to code complexity, manual memory management, and the limitations of

When Summaries Distort Decisions: Information Fidelity in LLM-Compressed Financial Analysis

AgentsDGX agent

arXiv:2606.29251v1 Announce Type: new Abstract: Financial decision-makers face more information than they can directly inspect, making context compression necessary. Yet when large language models (LL

29 Jun 2026

Cursor for iOS!

IndustryDGX agent

Cursor for iOS! Introducing Cursor for iOS. Build from anywhere by launching always-on cloud agents. Or remotely control agents running on your computer from the app. Composer 2.5 is 75% off in the ap

26 Jun 2026

Closing the Loop to Discover Psychological Theories with an Automated Cognitive Scientist

AgentsDGX agent

arXiv:2606.26448v1 Announce Type: cross Abstract: Across the sciences, autonomous systems are increasingly being used in closed-loop discovery, proposing new theories and designing and running experim

Content-Based Smart E-Mail Dispatcher Using Large Language Models

AgentsDGX agent

arXiv:2606.26593v1 Announce Type: new Abstract: Email communication has become an integral part of personal and professional life, but handling its vast volume is still a significant issue for large o

Finding the Time to Think: Learning Planning Budgets in Real-Time RL

SafetyDGX agent

arXiv:2606.26463v1 Announce Type: new Abstract: Deliberating takes time. In real-time settings, that time is not free. Standard reinforcement learning (RL) sidesteps this as the environment waits inde

From query to action: Introducing SQL alerting in Cloud Monitoring Observability Analytics

SafetyDGX agent

Traditional alerting systems often force a compromise: you can either alert immediately on simple, noisy log events, or monitor rigid, pre-configured metrics that fail when faced with data with many u

Life After Benchmark Saturation: A Case Study of CORE-Bench

Model ReleasesDGX agent

arXiv:2606.26158v1 Announce Type: new Abstract: When a benchmark's accuracy saturates, it is often retired and replaced with a more challenging version. We show that this approach privileges accuracy

OpenRCA 2.0: From Outcome Labels to Causal Process Supervision

Model ReleasesDGX agent

arXiv:2606.27154v1 Announce Type: new Abstract: Root cause analysis (RCA) poses a holistic test of LLM agentic capabilities, such as long-context understanding, multi-step reasoning, and tool use. How

The @n8n_io node for the LlamaParse Platform is now an officially verified community node🦙 Out of the box, you get access to parsing, split…

AgentsDGX agent

The @n8n_io node for the LlamaParse Platform is now an officially verified community node🦙 Out of the box, you get access to parsing, splitting, classification, structured data extraction and retrieva

25 Jun 2026

AI and Liability

ToolsDGX agent

AI and Liability Bruce Schneier on the recent German ruling that Google be held liable for errors introduced in their AI overviews: AI agents are agents of the person or organization that deploys them

daVinci-kernel: Co-Evolving Skill Selection, Summarization, and Utilization via RL for GPU Kernel Optimization

SafetyDGX agent

arXiv:2606.16497v2 Announce Type: replace-cross Abstract: GPU kernel optimization represents a paradigm where functional correctness is assumed and execution efficiency is the objective. We present da

SurveilNav: Collaborative Object Goal Navigation with Robot and Surveillance System

AgentsDGX agent

arXiv:2606.25119v1 Announce Type: new Abstract: With the growing deployment of surveillance systems in factories, offices, and homes, integrating them with robots offers a promising direction for coll

VADAOrchestra: Neurosymbolic Orchestration of Adaptive Reasoning Workflows

AgentsDGX agent

arXiv:2606.22485v2 Announce Type: replace-cross Abstract: Decision-making in real-world settings rarely follows a fixed script. Instead, it unfolds as a dynamic reasoning process in which the appropri

24 Jun 2026

Cryptographic certificates of validity for trustworthy AI

SafetyDGX agent

arXiv:2606.23768v1 Announce Type: cross Abstract: We propose cryptographic certificates of validity for agentic AI systems. The core idea is to formally specify a correctness or policy condition as a

LemonHarness Technical Report

Model ReleasesDGX agent

arXiv:2606.24311v1 Announce Type: new Abstract: As large language model (LLM) agents are applied to longer tasks, they increasingly modify workspace state across multiple rounds of iteration. However,

LOTS of alpha in this pod: - Why Databricks beat Snowflake (! a straight answer!) - Why everyone is building a metaharness now - Why the @ne…

AgentsDGX agent

LOTS of alpha in this pod: - Why Databricks beat Snowflake (! a straight answer!) - Why everyone is building a metaharness now - Why the @neondatabase made so much sense (so much @nikitabase glazing i

Managing Task Execution for Unknown Workloads in Batteryless IoT: A Hardware-Agnostic Evaluation

AgentsDGX agent

arXiv:2606.24340v1 Announce Type: new Abstract: In recent years, the Internet of Things (IoT) paradigm has been shifting toward batteryless, energy-harvesting architectures. Sustaining reliable operat

Toward Self-Evolution-Ready Workflow Harnesses: A Reversible Migration Path and Convertibility Taxonomy for Expert LLM Pipelines

AgentsDGX agent

arXiv:2606.24598v1 Announce Type: cross Abstract: While expert-validated 'LLM + script' workflows deliver significant value, they remain static: they encode hard-won domain knowledge yet fail to adapt

23 Jun 2026

interesting point here: loops amplify behavior, making them a double edged sword but we know loops are the future, so how do we avoid amplif…

AgentsDGX agent

interesting point here: loops amplify behavior, making them a double edged sword but we know loops are the future, so how do we avoid amplifying bad patterns? you need an *engaged* human in the (stack

OmniV2X: A Generative Foundation Planner for Efficient End-to-End Cooperative Driving

AgentsDGX agent

arXiv:2606.21165v1 Announce Type: new Abstract: We present OmniV2X, a generative foundation model for vehicle-to-everything (V2X) cooperative driving. The model directly interprets independent context

VTOS: Learning to Orchestrate Vision Tools by Co-Searching Solutions and Observers

AgentsDGX agent

arXiv:2606.20728v1 Announce Type: new Abstract: Vision foundation tools such as open-vocabulary detectors, segmentation models, and post-processing operators are powerful building blocks for computer

11 Jun 2026

ATLAS: Active Theory Learning for Automated Science

AgentsDGX agent

arXiv:2606.12386v1 Announce Type: cross Abstract: Advancing scientific understanding through mechanistic modeling requires posing the right experimental questions to yield maximally informative data.

← Previous
1…175176177178179…300
Next →