AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
Human
86,457Total entries
1Added by human
86,456Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
86,456 results
15 Jul 2026

Tracing Agentic Failure from the Flow of Success

AgentsDGX agent

arXiv:2607.12747v1 Announce Type: new Abstract: Failure attribution for LLM-based agentic systems, i.e., identifying which steps in a failure trajectory caused the task to fail, is critical for debugg

Track, Rank, Crack: Epistemic Working Memory Scales Multi-Hop Reasoning in Language Agents

Model ReleasesDGX agent

arXiv:2607.12267v1 Announce Type: cross Abstract: Language agents that interleave reasoning and tool use degrade sharply as reasoning chains lengthen, even when each individual step is easy. We trace

TRAIL: A Platform for Configurable Human--AI Teaming Experiments

SafetyDGX agent

arXiv:2607.12180v1 Announce Type: cross Abstract: An AI teammate's design properties (personality, communication style, when it speaks) can shape a team's trust, coordination, and decisions. Studying

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Training against GPT‑Red makes GPT‑5.6 substantially more resilient. To measure this, we replayed some of GPT‑Red’s strongest attacks—none o…

ApplicationsDGX agent

Training against GPT‑Red makes GPT‑5.6 substantially more resilient. To measure this, we replayed some of GPT‑Red’s strongest attacks—none of which our models had seen during training. GPT‑5.6 Sol pro

Traj-VLN: Learning Pixel-Space Interaction via Autoregressive Trajectory Generation

AgentsDGX agent

arXiv:2607.10744v2 Announce Type: replace Abstract: Benefiting from the powerful priors embedded in large-scale pre-training data and the emerging commonsense reasoning ability, large language models

Transforming LLMs into Efficient Cross-Encoders via Knowledge Distillation for RAG Reranking

Model ReleasesDGX agent

arXiv:2607.11933v1 Announce Type: new Abstract: Cross-encoders achieve high reranking accuracy in Retrieval-Augmented Generation (RAG) pipelines but impose quadratic inference costs that limit real-ti

Translation as a Computationally Efficient Bridge: Feasibility of English BERT for Low-Resource Languages

ResearchDGX agent

arXiv:2607.12612v1 Announce Type: new Abstract: BERT models have revolutionised Natural Language Processing (NLP) through their ability to process unstructured text across diverse domains. However, de

Trending AND fastest-growing in the same month? Benchmarks change daily, but only @tryramp has the database of real receipts to track this. …

Model ReleasesDGX agent

Trending AND fastest-growing in the same month? Benchmarks change daily, but only @tryramp has the database of real receipts to track this. We can confirm: demand for open-weight inference and trainin

TrustVLA: Mechanism-Guided Inference-Time Defense Against Vision-Language-Action Backdoors

SafetyDGX agent

arXiv:2607.12571v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models are deployed through pipelines that end users cannot audit, and a poisoned VLA can behave normally on clean observat

TSCA-Net: Temporal-Spatial Clique Attention for Interpretable Multimodal Pedestrian Trajectory Prediction

Local AiDGX agent

arXiv:2607.11939v1 Announce Type: new Abstract: Accurate pedestrian trajectory prediction in crowded environments remains challenging due to the multimodal uncertainty of human motion and the variable

Tune in today at 10am PST!

Model ReleasesDGX agent

Tune in today at 10am PST! Livestream Alert: Run ComfyUI From Claude/Cursor with Comfy MCP Host: @PurzBeats Comfy MCP lets Claude, Cursor, Amp and almost any AI agent you're already using build, run,

UMSS: Towards Unsupervised Multi-modal Semantic Segmentation

TutorialsDGX agent

arXiv:2607.12372v1 Announce Type: new Abstract: Multimodal semantic segmentation (MSS) is essential for robust perception in complex environments, yet its potential remains largely untapped because of

Uncertainty-Aware Multi-Source Retinal Fluid Segmentation in OCT

ResearchDGX agent

arXiv:2607.12212v1 Announce Type: cross Abstract: Measuring retinal fluid from optical coherence tomography (OCT) drives treatment decisions in macular disease, but manual annotation is slow and segme

Understanding Sources of Demographic Predictability in Brain MRI via Disentangling Anatomy and Contrast

SafetyDGX agent

arXiv:2603.04113v2 Announce Type: replace-cross Abstract: Demographic attributes can be predicted from medical images, raising concerns about bias in clinical AI systems. In X-ray imaging, acquisition

Understanding Structured Health Data through Interaction-Aware Mixture-of-Experts

ResearchDGX agent

arXiv:2607.12255v1 Announce Type: new Abstract: We study interaction-aware mixture-of-experts for post-stroke rigidity prediction using multi-level views of structured health records. Despite minimal

UniMedSeg: Unified In-Context Learning for Multi-Paradigm 2D/3D Medical Image Segmentation

ResearchDGX agent

arXiv:2607.12896v1 Announce Type: new Abstract: Medical image segmentation foundation models are expected to generalize across diverse clinical scenarios, yet existing universal methods remain fragmen

UniVR: Thinking in Visual Space for Unified Visual Reasoning

Model ReleasesDGX agent

arXiv:2607.12800v1 Announce Type: new Abstract: Learning broad world knowledge directly from raw visual data is a fundamental capability of intelligence. We introduce UniVR, the first investigation in

Unveiling Complex Collective Behaviors from Simple Rewards

AgentsDGX agent

arXiv:2607.12861v1 Announce Type: cross Abstract: Multi-agent Reinforcement Learning (MARL) holds great potential for robot swarms, but the black-box nature of neural policies complicates strategic an

UR-VC: Unsupervised Robotic Value Correction for Time-Derived Progress Proxies

Local AiDGX agent

arXiv:2607.12892v1 Announce Type: cross Abstract: Modern robot learning systems increasingly rely on dense progress or value signals to evaluate intermediate states, guide policy learning, and detect

VanillaBench: The Hidden Accuracy Cost of Adversarial Robustness

Model ReleasesDGX agent

arXiv:2607.12545v1 Announce Type: cross Abstract: Adversarial robustness research has produced hundreds of defended models over the past decade, yet the literature almost universally reports robustnes

Variational Mixture of Graph Neural Experts for Alzheimer's Disease Recognition across Frequency Bands in EEG Brain Networks

TutorialsDGX agent

arXiv:2510.11917v2 Announce Type: replace Abstract: Dementia disorders such as Alzheimer's disease (AD) and frontotemporal dementia (FTD) exhibit overlapping electrophysiological signatures in EEG tha

Verifier-Based Reinforcement Fine-Tuning of Reasoning Models for Thermal Energy Storage Control

Model ReleasesDGX agent

arXiv:2607.12856v1 Announce Type: new Abstract: Buildings are expected to shift cooling loads in response to grid conditions. Thermal energy storage (TES) enables this shift, but scheduling it well re

Vertical Standardisation for High-Risk AI Systems under the EU AI Act: A Domain-Specific Framework for Algorithmic Hiring

SafetyDGX agent

arXiv:2607.12588v1 Announce Type: new Abstract: According to the recent European legislation, high-risk AI systems will have to adapt in order to comply with requirements related to specific areas, li

ViCo3D: Empowering LiDAR-based Collaborative 3D Object Detection with Vision Foundation Models

AgentsDGX agent

arXiv:2607.12959v1 Announce Type: new Abstract: LiDAR-based collaborative 3D perception in Vehicle-to-Everything (V2X) systems typically relies on fusing bird's-eye-view (BEV) features across agents.

ViHoRec: A Quality-Controlled Vietnamese Hotel Recommendation Dataset and Cold-Start Benchmark

Model ReleasesDGX agent

arXiv:2607.12946v1 Announce Type: cross Abstract: Recommender-system research for Vietnamese remains limited by the absence of a public, well-documented hotel interaction resource. Building such a res

Virtual Chromoendscopy with Tunable Visibility Enhancement

ResearchDGX agent

arXiv:2607.12416v1 Announce Type: new Abstract: Chromoendoscopy (CE) is a common clinical practice that sprays indigo carmine blue dye onto the gastric surface to improve the visibility of gastric les

VISA: VLM-Guided Instance Semantic Auditing for 3D Occupancy World Models

AgentsDGX agent

arXiv:2606.13460v2 Announce Type: replace Abstract: Semantic 3D occupancy provides a voxelized world state for autonomous driving and robot decision making, but object and rare-class errors can affect

VisCo: Leveraging Large Language Models as Intrinsic Encoders for Visual Token Compression

Model ReleasesDGX agent

arXiv:2607.12756v1 Announce Type: new Abstract: Vision-language models (VLMs) process large numbers of visual tokens, resulting in substantial inference latency and memory overhead. This has motivated

Vision-Based Dribbling for Humanoid Soccer via Privileged Representation Learning

SafetyDGX agent

arXiv:2607.12702v1 Announce Type: new Abstract: Recent advances in humanoid robotics have highlighted the importance of deployable loco-manipulation skills. Dribbling a soccer ball while evading activ

VistaVLA: Geometry- and Semantic-Aware 3D Gaussian-Grounded VLA for Robotic Manipulation

SafetyDGX agent

arXiv:2607.12356v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have emerged as a powerful end-to-end paradigm for robotic manipulation by mapping language instructions and 2D visu

Visual Access Boundaries in Vision-Language Model Reasoning

ResearchDGX agent

arXiv:2607.12815v1 Announce Type: new Abstract: Chain-of-Thought (CoT) prompting is widely used as a test-time scaling strategy for Vision-Language Models (VLMs), but it remains unclear what is extend

Visual Species Recognition with Large Multimodal Models as Post-Hoc Correctors

ResearchDGX agent

arXiv:2512.15748v2 Announce Type: replace-cross Abstract: Visual Species Recognition (VSR) is a fundamental task in scientific disciplines that require species-level identification, including ecology,

VQCSim: When Does Compile-Once Statevector Simulation Beat Generic Quantum Frameworks?

HardwareDGX agent

arXiv:2607.11985v1 Announce Type: cross Abstract: Hybrid quantum-classical machine learning workflows repeatedly evaluate many small parametrized circuits during training and model exploration. In thi

WanToFight: Real-Time Generative Game Engine for Multi-Player Combat Interaction

Local AiDGX agent

arXiv:2607.12592v1 Announce Type: new Abstract: We present WanToFight, a generative game engine that simulates real-time, two-player The King of Fighters '97 (KOF~'97) gameplay from keyboard input. Pr

Watermark Forensics for Generative Models: An Information-Theoretic Perspective

SafetyDGX agent

arXiv:2607.13003v1 Announce Type: cross Abstract: A watermark in a generative model's output is usually asked only whether a text is machine-made. The same mark can do more: attribute it to the user w

We Hebben Een Serieus Translatie: Modeling Intercomprehension as Probabilistic Inference

SafetyDGX agent

arXiv:2607.12169v1 Announce Type: new Abstract: Intercomprehension refers to partial intelligibility of an unfamiliar language (L2) by a speaker of a related language (L1). How is this zero-shot cross

We made Together GPU Clusters more reliable and easier to operate. Passive health checks, guided node repair, a rebuilt Slurm stack, better …

HardwareDGX agent

We made Together GPU Clusters more reliable and easier to operate. Passive health checks, guided node repair, a rebuilt Slurm stack, better cluster visibility, external OIDC, startup scripts, and acce

We scaled a robot model natively to 8,000 timesteps of context, 5 minutes worth of muscle memory, with constant inference cost. Robot polici…

ResearchDGX agent

We scaled a robot model natively to 8,000 timesteps of context, 5 minutes worth of muscle memory, with constant inference cost. Robot policies used to live their lives a few frames at a time (< 0.1 se

Weakly Supervised Spatio-Temporal Candidate Discovery of Dairy Farm Sites from Seasonal Satellite Imagery

ResearchDGX agent

arXiv:2607.12748v1 Announce Type: cross Abstract: Farm site discovery from satellite imagery is a spatiotemporal candidate ranking problem because farm evidence is distributed across pasture, field bo

We're part of the Amazon Web Services (AWS) AI Builder Lab in New York on Friday, July 24 - a Clash of Agents competition with OpenAI, LangC…

AgentsDGX agent

We're part of the Amazon Web Services (AWS) AI Builder Lab in New York on Friday, July 24 - a Clash of Agents competition with OpenAI, LangChain, HiddenLayer, Protopia AI, Fiddler AI, and Coder. One d

What Does a Temporal Benchmark Score Measure? Decomposing Channel Use in Video VLM Evaluation

Model ReleasesDGX agent

arXiv:2607.12304v1 Announce Type: new Abstract: A score on a temporal video question answering benchmark is meant to measure that a model has temporal understanding, but it conflates two questions. 1.

What Does Goodness Measure? A Likelihood-Ratio Account of Forward-Forward Learning

Local AiDGX agent

arXiv:2607.12501v1 Announce Type: new Abstract: The Forward-Forward (FF) algorithm trains each layer locally, so that a scalar goodness - the sum of squared activations - is high on real inputs and lo

What Makes a Representational Prior Work? Feature Families, Label-Free Invariances, and Critical Windows in Grokking

SafetyDGX agent

arXiv:2607.12735v1 Announce Type: new Abstract: Companion work showed the grokking delay is causally the time to form task-structured representations, injectable via a contrastive prior. Here we chara

When and Why Does Multi-Agent Debate Fail and Does It Really Underperform?

AgentsDGX agent

arXiv:2510.20963v2 Announce Type: replace Abstract: Multi-agent debate (MAD) was proposed as a promising approach for ensembling the wisdom of multiple large language models (LLMs) to improve reasonin

When Close Enough Is Not Enough: Autoregressive Drift in Quantum Circuit Synthesis

Model ReleasesDGX agent

arXiv:2607.12780v1 Announce Type: cross Abstract: Quantum circuit optimization for fault-tolerant computing requires exact functional equivalence while minimizing expensive non-Clifford resources such

When Directional Accuracy Lies: A Base-Rate-Honest Benchmark for LoRA-Adapted TimesFM on Equity Forecasting

Model ReleasesDGX agent

arXiv:2607.12248v1 Announce Type: cross Abstract: Large pretrained time-series models such as TimesFM are attractive for financial forecasting, but raw directional accuracy is a misleading scoreboard

When Does Reward Teach State? A Hidden-Automaton Instrument and the Group-Language Boundary

SafetyDGX agent

arXiv:2607.11953v1 Announce Type: new Abstract: Does a reinforcement-learning agent that earns high reward represent its task's latent state, or only a reward-correlated shortcut? The question is usua

Who Grades the Grader? Co-Evolving Evaluation Metrics and Skills for Self-Improving LLM Agents

SafetyDGX agent

arXiv:2607.12790v1 Announce Type: new Abstract: Self-evolving agent systems improve by creating, revising, and retiring their own skills, but every such loop rests on a hidden assumption: a reliable e

Who touches every token that flows through @anthropic? It’s not any one model, but it’s @katelyn_lesse, @angjiang and the platform team. A y…

Model ReleasesDGX agent

Who touches every token that flows through @anthropic? It’s not any one model, but it’s @katelyn_lesse, @angjiang and the platform team. A year ago, it was just a messages API. Today, their platform s

WikiSTAR: A System for Shedding Light on the Hidden History of Scientific Wikipedia Articles

Model ReleasesDGX agent

arXiv:2607.12441v1 Announce Type: new Abstract: Wikipedia plays a key role in shaping public understanding of science, and its openly accessible revision history is a unique record of how scientific k

Win by Silence: Deletion Non-Monotonicity, Autonomous Exploitation, and Typed-State Gating in LLM Plan Evaluation

AgentsDGX agent

arXiv:2607.12986v1 Announce Type: new Abstract: Plan evaluators can reward a strategic plan for becoming less explicit. This paper studies that failure in a staged expected-value scorer for LLM-genera

X-Lens: Real-Time Metric Depth Estimation with Heterogeneous Cameras

Model ReleasesDGX agent

arXiv:2607.12993v1 Announce Type: new Abstract: We present X-lens, a compact feed-forward model for metric depth estimation from a variable number of calibrated fisheye and pinhole views. To support r

xai-org/grok-build, now open source

Model ReleasesDGX agent

xai-org/grok-build, now open source xAI's grok CLI tool faced severe community backlash yesterday when it became apparent that running the command in a directory could upload that entire directory to

Xray-Visual Models: Scaling Vision models on Industry Scale Data

ApplicationsDGX agent

arXiv:2602.16918v2 Announce Type: replace-cross Abstract: We present Xray-Visual, a unified vision model architecture for large-scale image and video understanding trained on industry-scale social med

You can use OpenCode Desktop with Ollama! Try it with the top open models!

Local AiDGX agent

You can use OpenCode Desktop with Ollama! Try it with the top open models! Introducing Tabs OpenCode Desktop is now built around tabs. Start a new session in a tab, or open an existing session from an

You don’t have to wait. Merch inspired by research & deployment. Available until sold out. https://openai.com/supply/

Model ReleasesDGX agent

OpenAI announced the launch of limited‑edition merchandise inspired by its research and deployment work, available for purchase until sold out via https://openai.com/supply/. The tweet highlighted “yo

Your coding agent doesn't need to leave the terminal to use Pinecone. We now ship official plugins and skills for the agentic IDEs and CLIs …

Model ReleasesDGX agent

Your coding agent doesn't need to leave the terminal to use Pinecone. We now ship official plugins and skills for the agentic IDEs and CLIs you are already building in: Claude Code, Cursor, GitHub Cop

14 Jul 2026

1) If you haven't read AI as Normal Technology, these annotated slides are probably the easiest way to get a high-level overview. https://ww…

ResearchDGX agent

1) If you haven't read AI as Normal Technology, these annotated slides are probably the easiest way to get a high-level overview. https://www.cs.princeton.edu/~arvindn/talks/icml-2026-annotated-slides

2.5x increase in usage of our agentic products (codex and chatgpt work) in the last week! welcome.

AgentsDGX agent

On July 14, 2026, OpenAI CEO Sam Altman announced a 2.5‑fold increase in usage of the company’s agentic products—Codex and ChatGPT‑based tools—within the preceding week. The tweet, which received 703.

🆕 5 Trends That Defined AI Engineering at World’s Fair 2026 https://latent.space/p/aiewf26trends @ricmac's big recap of @aidotengineer: 1. …

AgentsDGX agent

🆕 5 Trends That Defined AI Engineering at World’s Fair 2026 https://latent.space/p/aiewf26trends @ricmac's big recap of @aidotengineer: 1. The focus shifts from agents to systems 2. Loop engineering i

← Previous
1…281282283284285…1441
Next →