AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,343
  • Agents7,552
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,164
  • Local Ai4,928
  • Model Releases23,861
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,402

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,343
  • Agents7,552
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,164
  • Local Ai4,928
  • Model Releases23,861
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,402

Source
HumanDGX agent

88,343Total entries
1Added by human
88,342Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,572 results
10 Jul 2026

ZipDepth: Bringing Lightweight Zero-Shot Monocular Depth Anywhere, on Any Device

ResearchDGX agent

arXiv:2607.08771v1 Announce Type: new Abstract: Monocular depth estimation has seen remarkable progress through foundation models achieving robust zero-shot generalization, yet their computational dem

9 Jul 2026

A Gold-Standard Study of What Makes a Lightweight Game-Playing Agent Strong

Model ReleasesDGX agent

arXiv:2607.06854v1 Announce Type: cross Abstract: Reinforcement learning agents for imperfect-information card games are only as strong as the opponents they train against, and they are hard to grade,

A Quiet Failure in Calibrated Virtual Screening: Marginal Conformal Prediction Under-Covers the Minority Class, and a Class-Conditional Fix Recovers It

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Local AiDGX agent

arXiv:2607.06605v1 Announce Type: new Abstract: Conformal prediction is being adopted in drug discovery to put an honest number on model reliability: pick an error rate alpha, and the method returns p

Ace! Motion Planning of Professional-Level Table Tennis Serves with a Robot Arm

Model ReleasesDGX agent

arXiv:2607.06989v1 Announce Type: new Abstract: Table tennis, a dynamic, compact, and popular sport, has received significant attention as a robotics benchmark over the last decades. Most of the resea

AgentLens: Production-Assessed Trajectory Reviews for Coding Agent Evaluation

Model ReleasesDGX agent

arXiv:2607.06624v1 Announce Type: new Abstract: We present AgentLens, a production-assessed benchmark for interactive code agents. Most code-agent benchmarks reduce a run to a single bit -- did the ta

AI Chatbot Suicide Risk Detection and Response: Human Validation Study of the Open-Source VERA-MH Safety Evaluation

Model ReleasesDGX agent

arXiv:2602.05088v4 Announce Type: replace Abstract: Millions of people now use generative AI chatbots for psychological support. Despite their promise, the most pressing question in AI for mental heal

Best-Arm Identification with Generative Proxy

ResearchDGX agent

arXiv:2607.06879v1 Announce Type: new Abstract: Best-arm identification is a canonical model for data-driven decision-making, but in many applications each reward observation is costly. Motivated by t

Beyond Attack-Success Rate: Action-Graded Severity Scale for Tool-Using AI Agents

Model ReleasesDGX agent

arXiv:2607.07474v1 Announce Type: cross Abstract: Agentic red-teaming benchmarks report whether an injected agent was compromised as a single bit: the attack succeeded, or it did not. We argue that th

CaLiSym: Learning Symplectic Dynamics of Real-World Systems through Structured Canonical Lifts

Model ReleasesDGX agent

arXiv:2607.06824v1 Announce Type: cross Abstract: Physics-informed learning promises data-efficient and stable dynamics prediction, yet its strongest geometric guarantees have largely remained confine

Context-Aware Force Estimation for Deformable Tool Manipulation in Robotic Environmental Swabbing via Few-Shot Continual Adaptation

Model ReleasesDGX agent

arXiv:2607.07574v1 Announce Type: new Abstract: Robotic surface swabbing requires sustained interaction between a compliant tool and heterogeneous environments, where accurate estimation of tip-level

DiPhon: Diffusion on Graphons for Scalable Graph Generation

ResearchDGX agent

arXiv:2607.07232v1 Announce Type: cross Abstract: Diffusion models represent a leading paradigm for graph generation, with notable impact in domains such as molecular design. Yet, scaling these models

Dynamic-in-Few-Step: Unifying Dynamic Computation and Few-Step Distillation for Efficient Video Generation

ApplicationsDGX agent

arXiv:2607.06631v1 Announce Type: cross Abstract: Video Diffusion Models (VDMs) have demonstrated superior generation quality but suffer from prohibitive computational costs. While recent few-step dis

Ensemble Deep Learning Approaches for AI-Altered Video Detection

TutorialsDGX agent

arXiv:2607.06872v1 Announce Type: new Abstract: The increasing accessibility of artificial intelligence has led to a rapid rise in AI-generated videos, making it more difficult to distinguish between

EvoPlan: Evolutionary Neuro-Symbolic Robot Planning with Spatio-Temporal Guarantees

Model ReleasesDGX agent

arXiv:2607.06724v1 Announce Type: new Abstract: LLM-based robot planners are fluent but cannot guarantee that their plans are executable or safe. Classical PDDL planners can guarantee these properties

Explain Before You Answer: A Survey on Compositional Visual Reasoning

Model ReleasesDGX agent

arXiv:2508.17298v3 Announce Type: replace-cross Abstract: Compositional visual reasoning has emerged as a key research frontier in multimodal AI, aiming to endow machines with the human-like ability t

Grok 4.5

IndustryDGX agent

Grok 4.5 is an AI model developed by xAI, Elon Musk's artificial intelligence company, announced via his X platform. The model represents an advancement in xAI's Grok AI assistant line, likely featuri

Harrison Chase @hwchase17: Nemotron 3 Ultra hit 86%. Claude Opus hit 87%. At one-tenth the cost. Chase runs LangChain and disclosed it on st…

Model ReleasesDGX agent

Harrison Chase @hwchase17: Nemotron 3 Ultra hit 86%. Claude Opus hit 87%. At one-tenth the cost. Chase runs LangChain and disclosed it on stage: inside LangChain's internal deep-agents benchmark, open

Higher-Order Geometric Updates for Levenberg-Marquardt Method via Riemann Normal Coordinates

Model ReleasesDGX agent

arXiv:2607.07623v1 Announce Type: new Abstract: Nonlinear least-squares optimization is central to regression, physics-informed neural networks, and other machine-learning tasks. Such problems have a

InductWave: Inductive Multi-Hop Logical Query Answering on Knowledge Graphs

SafetyDGX agent

arXiv:2607.07422v1 Announce Type: new Abstract: Logical Multi-Hop Query Answering over Knowledge Graphs (KGs) can be formulated as querying, with an implicit completeness assumption. Current works mai

Max Out GRPO Signal: Adaptive Trace Prefix Control for Hard Reasoning Problems

SafetyDGX agent

arXiv:2607.07674v1 Announce Type: cross Abstract: Group Relative Policy Optimization (GRPO) stalls on a model's hardest problems: when no rollout in a group succeeds, the group-relative advantages van

MIRA-Math: A Benchmark for Minimal Information Requesting and Mathematical Reasoning

Model ReleasesDGX agent

arXiv:2607.07391v1 Announce Type: new Abstract: Mathematical reasoning benchmarks typically provide all facts needed to solve each problem, while interactive benchmarks often mix reasoning with tools,

Navigating Hierarchy: Hyperbolic Learning on Brain Graphs for Disorder Diagnosis

Local AiDGX agent

arXiv:2607.07077v1 Announce Type: cross Abstract: Functional brain networks exhibit a hierarchical organization across ROI, community, and whole-brain levels, supporting local processing, inter-commun

OpenAI rolls out GPT-5.6 after government greenlight — and announces ‘ChatGPT Work’

Model ReleasesDGX agent

About two weeks after OpenAI's GPT-5.6 was caught up in regulatory drama - rolled out only to government-approved organizations during a 'limited preview' period - the company has received the Trump a

Safely run AI-generated code in Cloud Run sandboxes

Model ReleasesDGX agent

Here’s a question we hear often at Google Cloud: How do you safely run AI-generated code or untrusted binaries without putting your host application, data, and cloud credentials at risk? In other word

Selective Timestep Weighting and Advantage-Based Replay for Sample-Efficient Diffusion RLHF

SafetyDGX agent

arXiv:2607.07693v1 Announce Type: cross Abstract: Reinforcement learning from human feedback (RLHF) has emerged as a powerful paradigm for aligning generative models with human preferences. However, a

SensorFM: Towards a general intelligence and interface for wearable health data

ResearchDGX agent

SensorFM is a foundation model that analyzes multimodal wearable sensor signals from devices like smartwatches and fitness trackers to generate insights into health and activities. The model was train

Smooth Operator: A Real-Time Sampling-Based Algorithm for Kinematic Hand Retargeting

TutorialsDGX agent

arXiv:2607.07491v1 Announce Type: new Abstract: Advances in learning-based robotic manipulation, such as Vision-Language-Action (VLA) models and Video Action Models (VAMs), heavily rely on high-qualit

STST-JEPA: Shallow-Target Spatio-Temporal Joint Embedding Prediction Architecture For EEG Self-Supervised Learning

ResearchDGX agent

arXiv:2607.06629v1 Announce Type: new Abstract: Brain age -- the age inferred from a physiological recording -- is an emerging biomarker whose deviation from chronological age tracks neurological and

The Key to Going Linear: Analysis-Driven Transformer Linearization

Model ReleasesDGX agent

arXiv:2607.07706v1 Announce Type: new Abstract: The quadratic cost of causal self-attention severely bottlenecks long-context transformer inference. While numerous post hoc linearization pipelines exi

Toward Robust Open-set Adaptation: Synapse Consolidation Inspired by Rac1/MAPK Pathways

ResearchDGX agent

arXiv:2604.00533v2 Announce Type: replace Abstract: Large Language Models (LLMs) generalize across tasks through reusable representations and flexible reasoning, yet remain brittle in real deployment

Towards Understanding Steering Strength

TutorialsDGX agent

arXiv:2602.02712v2 Announce Type: replace-cross Abstract: A popular approach to post-training control of large language models (LLMs) is the steering of intermediate latent representations. Namely, id

Trees from Marginals: Autoregressive drafting with factorized priors

HardwareDGX agent

arXiv:2607.06763v1 Announce Type: cross Abstract: Speculative decoding greatly increases the interactivity of autoregressive language models by trading off computation for extra tokens generated in a

VOTE: Vision-Language-Action Optimization with Trajectory Ensemble Voting

ResearchDGX agent

arXiv:2507.05116v5 Announce Type: replace-cross Abstract: Recent large-scale Vision Language Action (VLA) models have shown superior performance in robotic manipulation tasks guided by natural languag

What Predicts Correctness in Text-to-SQL? A Selective-Prediction Study

Model ReleasesDGX agent

arXiv:2607.06799v1 Announce Type: cross Abstract: Evaluating uncertainty in AI-generated SQL queries requires estimating whether a query is correct, where correct means it executes to the same result

8 Jul 2026

AbICL: In-Context Learning for Antigen-Specific Antibody Affinity Ranking

Model ReleasesDGX agent

arXiv:2607.05846v1 Announce Type: cross Abstract: Accurate ranking of antibody candidates according to their binding affinity is essential for therapeutic antibody discovery. However, existing methods

Back to the future

IndustryDGX agent

Back to the future Introducing... Hosted Models! 🌏 🏄‍♀️ Host Runway models online and connect to them anytime, anywhere, via a unique URL. Use them to create web pages, chatbots, plugins, and more. Th

Beyond Reactivity: Measuring Proactive Problem Solving in LLM Agents

Model ReleasesDGX agent

arXiv:2510.19771v4 Announce Type: replace Abstract: LLM-based agents are increasingly moving towards proactivity: rather than awaiting instruction, they exercise agency to anticipate user needs and so

CMDR: Contextual Multimodal Document Retrieval

Model ReleasesDGX agent

arXiv:2607.05927v1 Announce Type: cross Abstract: Multimodal document retrieval aims to retrieve relevant pages while preserving both textual and visual content from the original document. However, ex

CoPiT: Cognitive Pivot Translation for Digraphic Low-Resource Mongolian in the Traditional Script

Model ReleasesDGX agent

arXiv:2607.05849v1 Announce Type: new Abstract: Low-resource languages remain challenging for machine translation, and Mongolian is a representative case. As a digraphic language, Mongolian is written

Explainable embeddings with Distance Explainer

Model ReleasesDGX agent

arXiv:2505.15516v3 Announce Type: replace-cross Abstract: While eXplainable AI (XAI) has advanced significantly, few methods address interpretability in embedded vector spaces where dimensions represe

From Passive Observer to Active Critic: Reinforcement Learning Elicits Process Reasoning for Robotic Manipulation

Model ReleasesDGX agent

arXiv:2603.15600v2 Announce Type: replace-cross Abstract: Accurate process supervision remains a critical challenge for long-horizon robotic manipulation. A primary bottleneck is that current video ML

Great opportunity!

ApplicationsDGX agent

Great opportunity! We are hiring for @Harvey’s model training team. This team will help Harvey expand from the application layer into the model layer and from legal into high end knowledge work more b

Grok 4.5 in Grok Build also stands out for its efficiency. Grok 4.5 in Grok Build cost 2.49 per task while Fable 5 in Claude Code cost 11.…

Model ReleasesDGX agent

Grok 4.5 in Grok Build also stands out for its efficiency. Grok 4.5 in Grok Build cost 2.49 per task while Fable 5 in Claude Code cost 11.80 and GPT-5.5 in Codex $5.07. This is driven by relatively lo

How Personas Can Influence Agents to Play Split or Steal

TutorialsDGX agent

arXiv:2607.05398v1 Announce Type: new Abstract: Personas are often employed to guide large language model agents, yet their effectiveness in shaping strategic behavior in social dilemma settings remai

Improving TabPFN's Synthetic Data Generation by Integrating Causal Structure

Model ReleasesDGX agent

arXiv:2603.10254v2 Announce Type: replace Abstract: Synthetic tabular data generation addresses data scarcity and privacy constraints in a variety of domains. Tabular Prior-Data Fitted Network (TabPFN

IndoorR2X: Indoor Robot-to-Everything Coordination with LLM-Driven Planning

Model ReleasesDGX agent

arXiv:2603.20182v4 Announce Type: replace Abstract: Although robot-to-robot (R2R) communication improves indoor scene understanding beyond what a single robot can achieve, R2R alone cannot overcome pa

KernelEvolve: Scaling Agentic Kernel Coding for Heterogeneous AI Accelerators at Meta

SafetyDGX agent

arXiv:2512.23236v4 Announce Type: replace-cross Abstract: Making deep learning recommendation model (DLRM) training and inference fast and efficient is important. However, this presents three key syst

LingDT-VL-OCR: Structure-Aware Document-Level Parsing with Fine-Grained Visual Reference

Model ReleasesDGX agent

arXiv:2603.11044v2 Announce Type: replace Abstract: In this paper, we propose LingDT-VL-OCR, a document parsing system tailored to financial-domain documents, transforming ultra-long financial PDFs in

Nested Episodic State Topology (NEST): A Graph-Theoretic Architecture of Cognitive States

ResearchDGX agent

arXiv:2607.06055v1 Announce Type: cross Abstract: We present NEST (Nested Episodic State Topology), a foundational graph-theoretic representational ontology for modeling cognition as structured state

OrchardBench: A Physically-Grounded, GPU-Parallel Apple-Orchard Simulation Benchmark for Agricultural Robotics

Model ReleasesDGX agent

arXiv:2607.06337v1 Announce Type: cross Abstract: Robotic tree-fruit harvesting is a flagship problem for agricultural automation, but progress is bottlenecked by the cost and irreproducibility of fie

Regularity and Stability Properties of Selective SSMs with Discontinuous Gating

ResearchDGX agent

arXiv:2505.11602v3 Announce Type: replace Abstract: Selective State-Space Models (SSMs) such as Mamba have become central to long-sequence modeling. Still, their stability is poorly understood: their

Rewriting Bun in Rust

Model ReleasesDGX agent

Rewriting Bun in Rust Jarred Sumner has been promising this blog post (since May 9th) about his Zig to Rust rewrite of Bun for significantly longer than it took him to finish the rewrite. Honestly, it

The Balkanization of Execution-Security Research for AI Coding Agents: Isolation, Access Control, and Time-of-Check-to-Time-of-Use Vulnerabilities

Model ReleasesDGX agent

arXiv:2607.05743v1 Announce Type: cross Abstract: AI coding agents now read repositories, call tools, and execute shell commands with limited human oversight, and a fast-growing body of work studies w

Think Before You Grid-Search: Floor-First Triage for LLM Serving

Model ReleasesDGX agent

arXiv:2607.05876v1 Announce Type: cross Abstract: LLM serving optimization typically benchmarks many configurations and reaches for heavy profilers when latency targets are missed. We argue for the re

Vision as Unified Multimodal Generation

ResearchDGX agent

arXiv:2607.06560v1 Announce Type: new Abstract: We formulate computer vision as unified multimodal generation, where heterogeneous visual tasks are expressed in the native text and image generation sp

We raised 130M @ 1B for our series A To build the open superintelligence stack for everyone Pre-training concentrated frontier AI in a han…

HardwareDGX agent

We raised 130M @ 1B for our series A To build the open superintelligence stack for everyone Pre-training concentrated frontier AI in a handful of labs. RL changes who can build frontier AI and just wo

Where to cut, how deep: BPE and Unigram-LM on chemistry SMILES

ResearchDGX agent

arXiv:2607.05691v1 Announce Type: new Abstract: Every chemical language model reading SMILES begins with a tokenizer, yet the field has inherited byte-pair encoding (BPE) from natural language with li

7 Jul 2026

A Retrieval-Augmented Framework for Detecting and Resolving Pragmatic Ambiguities in Natural Language Requirements

Model ReleasesDGX agent

arXiv:2607.04436v1 Announce Type: cross Abstract: Natural language requirements (NLRs) are essential for bridging communication gaps among diverse stakeholders in software development. However, the in

A Step Towards Robust Unsupervised Domain Adaptation via Fine-Tuning and Reinforcement Learning

Model ReleasesDGX agent

arXiv:2607.03600v1 Announce Type: cross Abstract: Adversarial robustness in Unsupervised Domain Adaptation (UDA) remains a significant challenge due to noisy pseudo labels and inherent distributional

AgentLTL: A Trace-Verification Framework for Measuring, Enforcing, and Training Procedural Compliance in Tool-Using LLM Agents

Model ReleasesDGX agent

arXiv:2607.02599v1 Announce Type: cross Abstract: Tool-using LLM agents are usually evaluated by final-answer correctness or LLM judges. Neither captures how an answer was produced. In safety-critical

← Previous
1…452453454455456…1060
Next →