AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,724 results
7 Jul 2026

ResearchStudio-Reel: Automate the Last Mile of Research from Paper to Poster, Video, and Blog

Model ReleasesDGX agent

arXiv:2607.04438v1 Announce Type: cross Abstract: Research dissemination, turning a paper into a poster, a talk video, and a blog post, is still a manual last mile. Prior automation treats each artifa

Responsibility Distribution Estimation in Ego-View Accident Videos with Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2607.03591v1 Announce Type: cross Abstract: Recent studies on multimodal traffic accident understanding have mainly relied on infrastructure-camera footage, satellite imagery, or structured cras

RL-Ballast: Ship Ballast Water Path Planning and Clog Prediction via Reinforcement Learning

SafetyDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.04906v1 Announce Type: new Abstract: Under the Shipping 4.0 paradigm, autonomous and reduced-crew vessels require intelligent internal systems to maintain operational safety and structural

S-EMBER: A Large-Scale Benchmark for Streaming Egocentric Memory Retrieval

Model ReleasesDGX agent

arXiv:2607.02689v1 Announce Type: cross Abstract: As wearable devices enable continuous first-person recording, AI assistants must reason across long time horizons to recall past experiences-a capabil

Seduced by the Narrative: Assessing Rule Adherence in Semi-Open Textual Sandboxes

Model ReleasesDGX agent

arXiv:2607.02802v1 Announce Type: cross Abstract: As LLMs are increasingly deployed as autonomous adjudicators in semi-open textual game environments, robust rule adherence becomes critical when user

Spectral Rewiring for Exploration, Purification, and Model Merging

Model ReleasesDGX agent

arXiv:2607.03065v1 Announce Type: cross Abstract: Reinforcement learning has become a standard post-training recipe for large language models, but dense full-parameter updates create two deployment-re

STRATOS: Bridging the Symbolic-to-Numeric Gap in Spatio-Temporal Text-to-SQL for Meteorological Data

SafetyDGX agent

arXiv:2607.03501v1 Announce Type: cross Abstract: Copernicus, the European Union's Earth observation program, produces petabytes of Earth observation and climate data, offering immense potential for r

The AI labs desperately need non-engineer Peters, Borises, and Thariqs. Most demos for 'business users' are about replying to emails or to S…

Model ReleasesDGX agent

The AI labs desperately need non-engineer Peters, Borises, and Thariqs. Most demos for 'business users' are about replying to emails or to Slack. And yes, that's helpful to manage the cacophonous hell

The ‘Ghost’ in the Database: Recovering Active ADFS Signing Keys via Machine DPAPI

SafetyDGX agent

Written by: Shebin Mathew Introduction The 'Golden SAML' technique, first described by CyberArk researchers in 2017, and further detailed by Mandiant researchers in 2021, remains one of the most effec

TREK: Distill to Explore, Reinforce to Refine

Model ReleasesDGX agent

arXiv:2607.05339v1 Announce Type: cross Abstract: Group Relative Policy Optimization (GRPO) is effective when the current policy already samples useful reasoning trajectories, but it stalls on hard pr

Turning Off-Policy Tokens On-Policy: A Plug-in Approach for Improving LLM Alignment

SafetyDGX agent

arXiv:2607.04728v1 Announce Type: cross Abstract: Reinforcement learning (RL) post-training for large language models (LLMs) follows a efficient paradigm of 'rollout then update', which inevitably res

Understanding electricity consumption behaviour through Inverse Reinforcement Learning

ResearchDGX agent

arXiv:2607.03176v1 Announce Type: new Abstract: Understanding how households consume electricity in response to socioeconomic and climatic drivers is important for decision-makers designing energy pol

Unified Audio Intelligence Without Regressing on Text Intelligence

Model ReleasesDGX agent

arXiv:2607.05196v1 Announce Type: cross Abstract: Audio intelligence involves understanding, reasoning about, and generating both audio and speech. In this work, we introduce Nemotron-Labs-Audex-30B-A

v0.31.2

Local AiDGX agent

Ollama v0.31.2-rc1 is a pre-release version released on July 6, 2026. This release includes CI improvements to avoid unbounded parallelism, fixes for CUDA toolkit lookup, updates to cloud documentatio

Wan-Streamer v0.2: Higher Resolution, Same Latency

HardwareDGX agent

arXiv:2607.04443v1 Announce Type: cross Abstract: We present Wan-Streamer v0.2, a latency-preserving upgrade of the native-streaming, end-to-end audio-visual interaction model. v0.2 keeps the v0.1 mod

6 Jul 2026

I talk to a lot of companies that still have active efforts to build GPTs. (It remains weird that OpenAI abandoned GPTs after rolling them o…

ApplicationsDGX agent

I talk to a lot of companies that still have active efforts to build GPTs. (It remains weird that OpenAI abandoned GPTs after rolling them out. They were the precursor to Skills & could have been a br

// In-context Retrieval at Million-token Scale // Great study providing better understanding of retrieval at million-token scale. They run t…

TutorialsDGX agent

// In-context Retrieval at Million-token Scale // Great study providing better understanding of retrieval at million-token scale. They run the first systematic study of in-context retrieval at the sca

// ReContext // Models now support 128K context windows and still fail to use evidence that is already in the prompt. Where is the gap? New …

TutorialsDGX agent

// ReContext // Models now support 128K context windows and still fail to use evidence that is already in the prompt. Where is the gap? New paper introduces ReContext, a training-free inference harnes

4 Jul 2026

Better Models: Worse Tools

Model ReleasesDGX agent

Better Models: Worse Tools Armin reports on a weird problem he ran into while hacking on Pi: The short version is that newer Claude models sometimes call Pi’s edit tool with extra, invented fields in

3 Jul 2026

Activation Steering for Aligned Open-ended Generation without Sacrificing Coherence

Model ReleasesDGX agent

arXiv:2604.08169v2 Announce Type: replace Abstract: Alignment in LLMs is more brittle than commonly assumed: misalignment can be induced by adversarial prompts, benign fine-tuning, emergent misalignme

AI Virtue: What is 'Good' Knowledge in the Age of Artificial Intelligence?

ResearchDGX agent

arXiv:2607.01776v1 Announce Type: cross Abstract: In the age of AI, what will be good knowledge? This article, which is accepted and forthcoming in a special issue of Modern Fiction Studies on 'Cultur

Another major Grok Build update just landed, packed with new features, extensive bug fixes, and meaningful performance improvements Release …

SafetyDGX agent

Another major Grok Build update just landed, packed with new features, extensive bug fixes, and meaningful performance improvements Release Notes: v0.2.84 — 2026-07-03 Features: • Announcements now up

HaloGuard 1.0: An Open Weights Constitutional Classifier for Multilingual AI Safety

Model ReleasesDGX agent

arXiv:2607.02079v1 Announce Type: new Abstract: We present HaloGuard 1.0, an open-weights implementation of the constitutional-classifier paradigm for input safety. It achieves state-of-the-art perfor

Highly-recommended read from MIT on the part of RL with verifiable rewards that everyone keeps hitting. RLVR only optimizes what you can obj…

Model ReleasesDGX agent

Highly-recommended read from MIT on the part of RL with verifiable rewards that everyone keeps hitting. RLVR only optimizes what you can objectively score, so style, structure, and diversity quietly c

MMBench-Live: A Continuously Evolving Benchmark for Multimodal Models

Model ReleasesDGX agent

arXiv:2607.01813v1 Announce Type: cross Abstract: Evaluation benchmarks are essential for assessing vision-language models (VLMs), but most multimodal benchmarks are static, making them vulnerable to

NEW paper worth reading. (bookmark it) The basic idea is to pair a compressive recurrent state with a small exact memory, which helps to rec…

TutorialsDGX agent

NEW paper worth reading. (bookmark it) The basic idea is to pair a compressive recurrent state with a small exact memory, which helps to recover long-range recall without giving up the efficiency of l

Overthink-Triggered Slowdown Attacks on LVLM-Based Robotic Systems

SafetyDGX agent

arXiv:2607.01518v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) have been increasingly integrated into robotic systems. However, these models may exhibit overthinking behaviors,

Playing 20 Question Game with Policy-Based Reinforcement Learning

SafetyDGX agent

arXiv:1808.07645v5 Announce Type: replace-cross Abstract: The 20 Questions (Q20) game is a well known game which encourages deductive reasoning and creativity. In the game, the answerer first thinks o

PreScience: A Dataset and Benchmark for Scientific Forecasting

Model ReleasesDGX agent

arXiv:2602.20459v2 Announce Type: replace Abstract: Can AI systems trained on the existing scientific record forecast the advances that will follow? We introduce PreScience, a dataset and benchmark fo

StatEval: A Comprehensive Benchmark for Large Language Models in Statistics

Model ReleasesDGX agent

arXiv:2510.09517v2 Announce Type: replace Abstract: Despite rapid advances in large language models (LLMs), statistical reasoning remains underrepresented in existing LLM benchmarks, which often do no

This is true… but maybe less important than the fact that people don’t try ambitious things with these systems. Many models are excellent as…

Model ReleasesDGX agent

This is true… but maybe less important than the fact that people don’t try ambitious things with these systems. Many models are excellent as a Google replacement, for homework “help,” etc. It is someo

2 Jul 2026

AGC-Bench: Measuring Artificial General Creativity

Model ReleasesDGX agent

arXiv:2607.01152v1 Announce Type: new Abstract: Creativity research has debated whether creativity is domain-specific (e.g., visual, writing, science), and if it is psychometrically separable from gen

Another fascinating paper on LLM Judges. (bookmark it) It's from Amazon, and they show that if you run panels of LLM judges, averaging their…

Model ReleasesDGX agent

Another fascinating paper on LLM Judges. (bookmark it) It's from Amazon, and they show that if you run panels of LLM judges, averaging their scores is a trap. 'Overall, we establish that robust aggreg

Beyond Document Grounding: Span-Level Hallucination Detection over Code, Tool Output, and Documents

Model ReleasesDGX agent

arXiv:2607.00895v1 Announce Type: new Abstract: Hallucination detection for retrieval-augmented generation (RAG) is usually evaluated on natural-language document evidence. However, grounded generatio

Bounded Morality: Defining the Space of Moral Computation

SafetyDGX agent

arXiv:2607.00002v1 Announce Type: new Abstract: Moral cognition has traditionally been modeled as adherence to fixed ethical theories--deontology, consequentialism, virtue ethics--implemented as stati

Distributed Multi Robot Lunar Cargo Transportation via Phase Decomposed Reinforcement Learning

SafetyDGX agent

arXiv:2607.00160v1 Announce Type: new Abstract: Modular reconfigurable robotic systems provide a scalable solution for cooperative surface operations in future lunar missions. However, cooperative car

ECoSim: Data Efficient Fine-Tuning for Controllable Traffic Simulation

SafetyDGX agent

arXiv:2607.00545v1 Announce Type: new Abstract: Controllable traffic simulation is critical for testing autonomous driving systems, yet existing approaches often require retraining large generative mo

FLYNN: Robust Neural Network for Robot Navigation using Fly Brain Topology

Model ReleasesDGX agent

arXiv:2607.00025v1 Announce Type: cross Abstract: While deep learning models achieve state-of-the-art performance in complex tasks, they remain brittle when faced with new environments or sensory depr

I just left the final day of the @aiDotEngineer World's Fair Conference in San Francisco. Kudos to @swyx for putting together a world-class …

Model ReleasesDGX agent

I just left the final day of the @aiDotEngineer World's Fair Conference in San Francisco. Kudos to @swyx for putting together a world-class lineup of speakers and workshops! It really was an invigorat

Interact3D: Compositional 3D Generation of Interactive Objects

Local AiDGX agent

arXiv:2603.16085v2 Announce Type: replace-cross Abstract: Recent breakthroughs in 3D generation have enabled the synthesis of high-fidelity individual assets. However, generating 3D compositional obje

Is One Layer Enough? Training A Single Transformer Layer Can Match Full-Parameter RL Training

Model ReleasesDGX agent

arXiv:2607.01232v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become a central component of post-training large language models (LLMs), yet little is understood about how RL adapta

Learning Gait-Aware Quadruped Locomotion with Temporal Logic Specifications

SafetyDGX agent

arXiv:2607.00442v1 Announce Type: cross Abstract: Reinforcement learning (RL) for quadruped locomotion commonly depends on fixed, hand-crafted, and Markovian reward functions that limit both interpret

Learning Generalizable Skill Policy with Data-Efficient Unsupervised RL

SafetyDGX agent

arXiv:2607.00392v1 Announce Type: cross Abstract: Unsupervised Reinforcement Learning (URL) aims to pre-train scalable, skill-conditioned policies without extrinsic rewards, serving as a foundation fo

Learning to Watch: Active Video Anomaly Understanding via Interleaved Policy Optimization

Local AiDGX agent

arXiv:2607.00622v1 Announce Type: new Abstract: Video anomaly understanding (VAU) relies on sparse, context-dependent cues. However, existing passive paradigms suffer from observational aliasing, wher

LLM-Guided ODE Discovery and Parameter Inference from Small-Cohort Aggregate Data

Model ReleasesDGX agent

arXiv:2607.00733v1 Announce Type: cross Abstract: Mechanistic modeling via ordinary differential equations (ODEs) provides interpretable descriptions of complex dynamics and enables inference of under

My one serious piece of advice having used Fable a bunch before release is that, unless you are careful it develops its own internal bizarre…

ApplicationsDGX agent

My one serious piece of advice having used Fable a bunch before release is that, unless you are careful it develops its own internal bizarre cadence & dialogue over long tasks. If you aren't asking it

NEW paper from NVIDIA. They discuss robot programming that compounds experience instead of throwing it away. Traditional robot programming f…

SafetyDGX agent

NEW paper from NVIDIA. They discuss robot programming that compounds experience instead of throwing it away. Traditional robot programming forces you to orchestrate perception, contact dynamics, diver

New research from Google. LLMs hallucinate with high confidence, miss their own knowledge boundaries, and misreport uncertainty. Most fixes …

TutorialsDGX agent

New research from Google. LLMs hallucinate with high confidence, miss their own knowledge boundaries, and misreport uncertainty. Most fixes bolt calibration on from the outside. RLMF turns the model o

so proud to be working with a bunch of people who are absolutely crushing it! check out all these new things if you haven’t had a chance to …

Model ReleasesDGX agent

so proud to be working with a bunch of people who are absolutely crushing it! check out all these new things if you haven’t had a chance to yet. huge week for us here @LangChain!! big week at langchai

Stop Pretending Social Robots Are Inevitable

ResearchDGX agent

arXiv:2607.00142v1 Announce Type: new Abstract: This paper takes issue with the recent themes of both the RO-MAN and the HRI conferences for their portrayal of a future human-robot society as inevitab

Task-Relevant Representation Decoupling for Visual Reinforcement Learning Generalization

ResearchDGX agent

arXiv:2607.00796v1 Announce Type: new Abstract: Visual Reinforcement Learning (VRL) has achieved considerable success in solving control tasks. However, generalizing learned policies to new environmen

🌍 World’s Largest Hermes Buildathon 10 cities. 5 countries. $750K credits Backed by OpenAI, Convex, Cloudflare, ElevenLabs, Linkup, Dodo Pa…

HardwareDGX agent

🌍 World’s Largest Hermes Buildathon 10 cities. 5 countries. $750K credits Backed by OpenAI, Convex, Cloudflare, ElevenLabs, Linkup, Dodo Payments, Hissa Fund & Wispr Flow. Proudly presented by GrowthX

1 Jul 2026

AC3S: Adaptive Conditioning for 3D-Aware Synthetic Data Generation

SafetyDGX agent

arXiv:2606.31204v1 Announce Type: new Abstract: Synthetic data generation has emerged as a powerful tool for improving data scalability in computer vision. Recent diffusion-based pipelines have demons

B2B sales workspace startup Aligned raised a 60M Series B led by PeakSpan Capital, taking its total funding to 73.8M, and says it has 1,000+ customers (Chris Metinko/Axios)

ApplicationsDGX agent

Chris Metinko / Axios: B2B sales workspace startup Aligned raised a 60M Series B led by PeakSpan Capital, taking its total funding to 73.8M, and says it has 1,000+ customers — Aligned, a sales workspa

Beyond Compilation: Evaluating Faithful Natural-Language-to-Lean Statement Formalization

Model ReleasesDGX agent

arXiv:2606.31002v1 Announce Type: new Abstract: Theorem-proving benchmarks evaluate proof search against fixed formal statements, but natural-language-to-Lean formalization must generate the formal st

Bridging Local Observation and Global Simulation in Closed-Loop Traffic Modeling

Local AiDGX agent

arXiv:2606.31844v1 Announce Type: cross Abstract: A local-to-global context mismatch arises when autoregressive traffic simulators trained on ego-centric driving logs are deployed in globally observab

Dataset Construction for Training LLM to Learn Analog Circuit Knowledge

Model ReleasesDGX agent

arXiv:2508.10409v3 Announce Type: replace-cross Abstract: This paper constructs a textual dataset for training large language models (LLMs) to learn analog circuit knowledge and customizes LLM trainin

ENPIRE -> ASPIRE, our 2nd work in the series for Physical AutoResearch. We are building the components for robot self-improvement, one /skil…

ResearchDGX agent

ENPIRE -> ASPIRE, our 2nd work in the series for Physical AutoResearch. We are building the components for robot self-improvement, one /skill at a time. Today, we give robots a /skills library that se

GaussLite: Online Task-Conditioned 3D Gaussian Splatting for Real-Time Robotic Mapping

ResearchDGX agent

arXiv:2606.30809v1 Announce Type: new Abstract: Existing 3D Gaussian Splatting (3DGS) systems distribute representation capacity uniformly across a scene, ignoring the fact that many downstream roboti

Knowledge Distillation from Large Reasoning Models to Compact Student Models: A Case Study on the John O Bryan Mathematics Competition

Model ReleasesDGX agent

arXiv:2606.31048v1 Announce Type: cross Abstract: This paper investigates knowledge distillation from a large reasoning model (DeepSeek-R1) to a compact student model (Qwen2.5-7B). Using historical pr

← Previous
1…269270271272273…296
Next →