AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,724 results
Model Releases

MMBench-Live: A Continuously Evolving Benchmark for Multimodal Models

DGX agent

arXiv:2607.01813v1 Announce Type: cross Abstract: Evaluation benchmarks are essential for assessing vision-language models (VLMs), but most multimodal benchmarks are static, making them vulnerable to

model-releasesarxiv-cs-ai
3 Jul 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Tutorials

NEW paper worth reading. (bookmark it) The basic idea is to pair a compressive recurrent state with a small exact memory, which helps to rec…

DGX agent

NEW paper worth reading. (bookmark it) The basic idea is to pair a compressive recurrent state with a small exact memory, which helps to recover long-range recall without giving up the efficiency of l

tutorialsdair-ai--x
3 Jul 2026
Safety

Overthink-Triggered Slowdown Attacks on LVLM-Based Robotic Systems

DGX agent

arXiv:2607.01518v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) have been increasingly integrated into robotic systems. However, these models may exhibit overthinking behaviors,

safetyarxiv-cs-ro
3 Jul 2026
Safety

Playing 20 Question Game with Policy-Based Reinforcement Learning

DGX agent

arXiv:1808.07645v5 Announce Type: replace-cross Abstract: The 20 Questions (Q20) game is a well known game which encourages deductive reasoning and creativity. In the game, the answerer first thinks o

safetyarxiv-cs-ai
3 Jul 2026
Model Releases

PreScience: A Dataset and Benchmark for Scientific Forecasting

DGX agent

arXiv:2602.20459v2 Announce Type: replace Abstract: Can AI systems trained on the existing scientific record forecast the advances that will follow? We introduce PreScience, a dataset and benchmark fo

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

StatEval: A Comprehensive Benchmark for Large Language Models in Statistics

DGX agent

arXiv:2510.09517v2 Announce Type: replace Abstract: Despite rapid advances in large language models (LLMs), statistical reasoning remains underrepresented in existing LLM benchmarks, which often do no

model-releasesarxiv-cs-cl
3 Jul 2026
Model Releases

This is true… but maybe less important than the fact that people don’t try ambitious things with these systems. Many models are excellent as…

DGX agent

This is true… but maybe less important than the fact that people don’t try ambitious things with these systems. Many models are excellent as a Google replacement, for homework “help,” etc. It is someo

model-releasesethan-mollick--x
3 Jul 2026
Model Releases

AGC-Bench: Measuring Artificial General Creativity

DGX agent

arXiv:2607.01152v1 Announce Type: new Abstract: Creativity research has debated whether creativity is domain-specific (e.g., visual, writing, science), and if it is psychometrically separable from gen

model-releasesarxiv-cs-cl
2 Jul 2026
Model Releases

Another fascinating paper on LLM Judges. (bookmark it) It's from Amazon, and they show that if you run panels of LLM judges, averaging their…

DGX agent

Another fascinating paper on LLM Judges. (bookmark it) It's from Amazon, and they show that if you run panels of LLM judges, averaging their scores is a trap. 'Overall, we establish that robust aggreg

model-releasesdair-ai--x
2 Jul 2026
Model Releases

Beyond Document Grounding: Span-Level Hallucination Detection over Code, Tool Output, and Documents

DGX agent

arXiv:2607.00895v1 Announce Type: new Abstract: Hallucination detection for retrieval-augmented generation (RAG) is usually evaluated on natural-language document evidence. However, grounded generatio

model-releasesarxiv-cs-cl
2 Jul 2026
Safety

Bounded Morality: Defining the Space of Moral Computation

DGX agent

arXiv:2607.00002v1 Announce Type: new Abstract: Moral cognition has traditionally been modeled as adherence to fixed ethical theories--deontology, consequentialism, virtue ethics--implemented as stati

safetyarxiv-cs-ai
2 Jul 2026
Safety

Distributed Multi Robot Lunar Cargo Transportation via Phase Decomposed Reinforcement Learning

DGX agent

arXiv:2607.00160v1 Announce Type: new Abstract: Modular reconfigurable robotic systems provide a scalable solution for cooperative surface operations in future lunar missions. However, cooperative car

safetyarxiv-cs-ro
2 Jul 2026
Safety

ECoSim: Data Efficient Fine-Tuning for Controllable Traffic Simulation

DGX agent

arXiv:2607.00545v1 Announce Type: new Abstract: Controllable traffic simulation is critical for testing autonomous driving systems, yet existing approaches often require retraining large generative mo

safetyarxiv-cs-cv
2 Jul 2026
Model Releases

FLYNN: Robust Neural Network for Robot Navigation using Fly Brain Topology

DGX agent

arXiv:2607.00025v1 Announce Type: cross Abstract: While deep learning models achieve state-of-the-art performance in complex tasks, they remain brittle when faced with new environments or sensory depr

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

I just left the final day of the @aiDotEngineer World's Fair Conference in San Francisco. Kudos to @swyx for putting together a world-class …

DGX agent

I just left the final day of the @aiDotEngineer World's Fair Conference in San Francisco. Kudos to @swyx for putting together a world-class lineup of speakers and workshops! It really was an invigorat

model-releasesswyx--x
2 Jul 2026
Local Ai

Interact3D: Compositional 3D Generation of Interactive Objects

DGX agent

arXiv:2603.16085v2 Announce Type: replace-cross Abstract: Recent breakthroughs in 3D generation have enabled the synthesis of high-fidelity individual assets. However, generating 3D compositional obje

local-aiarxiv-cs-ai
2 Jul 2026
Model Releases

Is One Layer Enough? Training A Single Transformer Layer Can Match Full-Parameter RL Training

DGX agent

arXiv:2607.01232v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become a central component of post-training large language models (LLMs), yet little is understood about how RL adapta

model-releasesarxiv-cs-cl
2 Jul 2026
Safety

Learning Gait-Aware Quadruped Locomotion with Temporal Logic Specifications

DGX agent

arXiv:2607.00442v1 Announce Type: cross Abstract: Reinforcement learning (RL) for quadruped locomotion commonly depends on fixed, hand-crafted, and Markovian reward functions that limit both interpret

safetyarxiv-cs-ai
2 Jul 2026
Safety

Learning Generalizable Skill Policy with Data-Efficient Unsupervised RL

DGX agent

arXiv:2607.00392v1 Announce Type: cross Abstract: Unsupervised Reinforcement Learning (URL) aims to pre-train scalable, skill-conditioned policies without extrinsic rewards, serving as a foundation fo

safetyarxiv-cs-ai
2 Jul 2026
Local Ai

Learning to Watch: Active Video Anomaly Understanding via Interleaved Policy Optimization

DGX agent

arXiv:2607.00622v1 Announce Type: new Abstract: Video anomaly understanding (VAU) relies on sparse, context-dependent cues. However, existing passive paradigms suffer from observational aliasing, wher

local-aiarxiv-cs-cv
2 Jul 2026
Model Releases

LLM-Guided ODE Discovery and Parameter Inference from Small-Cohort Aggregate Data

DGX agent

arXiv:2607.00733v1 Announce Type: cross Abstract: Mechanistic modeling via ordinary differential equations (ODEs) provides interpretable descriptions of complex dynamics and enables inference of under

model-releasesarxiv-cs-ai
2 Jul 2026
Applications

My one serious piece of advice having used Fable a bunch before release is that, unless you are careful it develops its own internal bizarre…

DGX agent

My one serious piece of advice having used Fable a bunch before release is that, unless you are careful it develops its own internal bizarre cadence & dialogue over long tasks. If you aren't asking it

applicationsethan-mollick--x
2 Jul 2026
Safety

NEW paper from NVIDIA. They discuss robot programming that compounds experience instead of throwing it away. Traditional robot programming f…

DGX agent

NEW paper from NVIDIA. They discuss robot programming that compounds experience instead of throwing it away. Traditional robot programming forces you to orchestrate perception, contact dynamics, diver

safetydair-ai--x
2 Jul 2026
Tutorials

New research from Google. LLMs hallucinate with high confidence, miss their own knowledge boundaries, and misreport uncertainty. Most fixes …

DGX agent

New research from Google. LLMs hallucinate with high confidence, miss their own knowledge boundaries, and misreport uncertainty. Most fixes bolt calibration on from the outside. RLMF turns the model o

tutorialsdair-ai--x
2 Jul 2026
Model Releases

so proud to be working with a bunch of people who are absolutely crushing it! check out all these new things if you haven’t had a chance to …

DGX agent

so proud to be working with a bunch of people who are absolutely crushing it! check out all these new things if you haven’t had a chance to yet. huge week for us here @LangChain!! big week at langchai

model-releasesharrison-chase--x
2 Jul 2026
Research

Stop Pretending Social Robots Are Inevitable

DGX agent

arXiv:2607.00142v1 Announce Type: new Abstract: This paper takes issue with the recent themes of both the RO-MAN and the HRI conferences for their portrayal of a future human-robot society as inevitab

researcharxiv-cs-ro
2 Jul 2026
Research

Task-Relevant Representation Decoupling for Visual Reinforcement Learning Generalization

DGX agent

arXiv:2607.00796v1 Announce Type: new Abstract: Visual Reinforcement Learning (VRL) has achieved considerable success in solving control tasks. However, generalizing learned policies to new environmen

researcharxiv-cs-lg
2 Jul 2026
Hardware

🌍 World’s Largest Hermes Buildathon 10 cities. 5 countries. $750K credits Backed by OpenAI, Convex, Cloudflare, ElevenLabs, Linkup, Dodo Pa…

DGX agent

🌍 World’s Largest Hermes Buildathon 10 cities. 5 countries. $750K credits Backed by OpenAI, Convex, Cloudflare, ElevenLabs, Linkup, Dodo Payments, Hissa Fund & Wispr Flow. Proudly presented by GrowthX

hardwarenous-research--x
2 Jul 2026
Safety

AC3S: Adaptive Conditioning for 3D-Aware Synthetic Data Generation

DGX agent

arXiv:2606.31204v1 Announce Type: new Abstract: Synthetic data generation has emerged as a powerful tool for improving data scalability in computer vision. Recent diffusion-based pipelines have demons

safetyarxiv-cs-cv
1 Jul 2026
Applications

B2B sales workspace startup Aligned raised a 60M Series B led by PeakSpan Capital, taking its total funding to 73.8M, and says it has 1,000+ customers (Chris Metinko/Axios)

DGX agent

Chris Metinko / Axios: B2B sales workspace startup Aligned raised a 60M Series B led by PeakSpan Capital, taking its total funding to 73.8M, and says it has 1,000+ customers — Aligned, a sales workspa

applicationstechmeme
1 Jul 2026
Model Releases

Beyond Compilation: Evaluating Faithful Natural-Language-to-Lean Statement Formalization

DGX agent

arXiv:2606.31002v1 Announce Type: new Abstract: Theorem-proving benchmarks evaluate proof search against fixed formal statements, but natural-language-to-Lean formalization must generate the formal st

model-releasesarxiv-cs-ai
1 Jul 2026
Local Ai

Bridging Local Observation and Global Simulation in Closed-Loop Traffic Modeling

DGX agent

arXiv:2606.31844v1 Announce Type: cross Abstract: A local-to-global context mismatch arises when autoregressive traffic simulators trained on ego-centric driving logs are deployed in globally observab

local-aiarxiv-cs-ai
1 Jul 2026
Model Releases

Dataset Construction for Training LLM to Learn Analog Circuit Knowledge

DGX agent

arXiv:2508.10409v3 Announce Type: replace-cross Abstract: This paper constructs a textual dataset for training large language models (LLMs) to learn analog circuit knowledge and customizes LLM trainin

model-releasesarxiv-cs-ai
1 Jul 2026
Research

ENPIRE -> ASPIRE, our 2nd work in the series for Physical AutoResearch. We are building the components for robot self-improvement, one /skil…

DGX agent

ENPIRE -> ASPIRE, our 2nd work in the series for Physical AutoResearch. We are building the components for robot self-improvement, one /skill at a time. Today, we give robots a /skills library that se

researchjim-fan--x
1 Jul 2026
Research

GaussLite: Online Task-Conditioned 3D Gaussian Splatting for Real-Time Robotic Mapping

DGX agent

arXiv:2606.30809v1 Announce Type: new Abstract: Existing 3D Gaussian Splatting (3DGS) systems distribute representation capacity uniformly across a scene, ignoring the fact that many downstream roboti

researcharxiv-cs-cv
1 Jul 2026
Model Releases

Knowledge Distillation from Large Reasoning Models to Compact Student Models: A Case Study on the John O Bryan Mathematics Competition

DGX agent

arXiv:2606.31048v1 Announce Type: cross Abstract: This paper investigates knowledge distillation from a large reasoning model (DeepSeek-R1) to a compact student model (Qwen2.5-7B). Using historical pr

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Scaling LLM Inference: Multi-Node KV Cache Offloading with GKE & Managed Lustre

DGX agent

Significant contributors to this article include Sneha Aradhey, Software Engineer, Google Kubernetes Engine, and Michael MacDonald, Sr Software Engineer, Google Cloud Managed Lustre. Enterprise produc

model-releasesgoogle-cloud-ai
1 Jul 2026
Safety

Stabilization Learning: A Paradigm Transition Bridging Control Theory and Machine Learning

DGX agent

arXiv:2606.31562v1 Announce Type: new Abstract: Stabilization learning is an interdisciplinary paradigm that bridges control theory and machine learning. Its core idea is to enable systems to adjust t

safetyarxiv-cs-ro
1 Jul 2026
Research

Stage-Transition Dense Reward Modeling for Reinforcement Learning

DGX agent

arXiv:2606.31377v1 Announce Type: cross Abstract: Reinforcement learning for long-horizon robotic manipulation is often limited by sparse and delayed rewards, while manually designing dense shaping si

researcharxiv-cs-ai
1 Jul 2026
Model Releases

Truth or Sophistry? LoFa: A Benchmark for LLM Robustness Against Logical Fallacies

DGX agent

arXiv:2606.31039v1 Announce Type: new Abstract: Large Language Models (LLMs) exhibit strong semantic capabilities, yet their resilience to manipulative linguistic patterns such as logical fallacies re

model-releasesarxiv-cs-cl
1 Jul 2026
Model Releases

What If We Allocate Test-Time Compute Adaptively?

DGX agent

arXiv:2602.01070v5 Announce Type: replace Abstract: Test-time compute scaling allocates inference computation uniformly, uses fixed sampling strategies, and applies verification only for reranking. In

model-releasesarxiv-cs-cl
1 Jul 2026
Model Releases

You really need to benchmark models for your use case. As soon as judgements & decisions stack on top of each other, the differences between…

DGX agent

You really need to benchmark models for your use case. As soon as judgements & decisions stack on top of each other, the differences between models amplifies, and no standard benchmark will tell you t

model-releasesethan-mollick--x
1 Jul 2026
Research

A Linear Matching Bandit Approach to Online Multi-Human Multi-Robot Teaming

DGX agent

arXiv:2606.29221v1 Announce Type: new Abstract: We address the problem of online multi-human multi-robot teaming through the lens of a linear matching bandit framework, where a learner assigns robots

researcharxiv-cs-lg
30 Jun 2026
Model Releases

A Machine-Verified Proof of a Quantum-Optimization Conjecture

DGX agent

arXiv:2606.29687v1 Announce Type: cross Abstract: We report a machine-verified resolution of a problem open for over a decade in quantum optimization: the Farhi, Goldstone and Gutmann (FGG) conjecture

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Conversational Query Engine for Mixed-Modality Heterogeneous Enterprise Data Sources

DGX agent

arXiv:2606.28370v1 Announce Type: cross Abstract: Enterprise business intelligence queries span structured warehouses and unstructured document repositories -- modalities with fundamentally different

model-releasesarxiv-cs-ai
30 Jun 2026
Research

Cross-Session 3D LiDAR and Camera Fusion for Robust Localization of Unmanned Aerial Vehicles in GPS-Denied Environments

DGX agent

arXiv:2606.28951v1 Announce Type: new Abstract: Accurate localization of unmanned aerial vehicles (UAVs) is essential for applications such as structural health monitoring, especially in environments

researcharxiv-cs-ro
30 Jun 2026
Research

Cybersecurity is the True Frontier for Generative AI Success or Failure

DGX agent

arXiv:2606.28929v1 Announce Type: cross Abstract: Cybersecurity is a real-life test-bed for many machine learning problems at once, especially when considering modern strides in using Large Language M

researcharxiv-cs-lg
30 Jun 2026
Model Releases

DiscoGen: Procedural Generation of Algorithm Discovery Tasks in Machine Learning

DGX agent

arXiv:2603.17863v2 Announce Type: replace-cross Abstract: Automating the development of machine learning algorithms has the potential to unlock new breakthroughs. However, our ability to improve and e

model-releasesarxiv-cs-ai
30 Jun 2026
← Previous
1…337338339340341…370
Next →