AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

88,419Total entries
1Added by human
88,418Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,638 results
9 Jul 2026

Smooth Operator: A Real-Time Sampling-Based Algorithm for Kinematic Hand Retargeting

TutorialsDGX agent

arXiv:2607.07491v1 Announce Type: new Abstract: Advances in learning-based robotic manipulation, such as Vision-Language-Action (VLA) models and Video Action Models (VAMs), heavily rely on high-qualit

STST-JEPA: Shallow-Target Spatio-Temporal Joint Embedding Prediction Architecture For EEG Self-Supervised Learning

ResearchDGX agent

arXiv:2607.06629v1 Announce Type: new Abstract: Brain age -- the age inferred from a physiological recording -- is an emerging biomarker whose deviation from chronological age tracks neurological and

The Key to Going Linear: Analysis-Driven Transformer Linearization

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.07706v1 Announce Type: new Abstract: The quadratic cost of causal self-attention severely bottlenecks long-context transformer inference. While numerous post hoc linearization pipelines exi

Toward Robust Open-set Adaptation: Synapse Consolidation Inspired by Rac1/MAPK Pathways

ResearchDGX agent

arXiv:2604.00533v2 Announce Type: replace Abstract: Large Language Models (LLMs) generalize across tasks through reusable representations and flexible reasoning, yet remain brittle in real deployment

Towards Understanding Steering Strength

TutorialsDGX agent

arXiv:2602.02712v2 Announce Type: replace-cross Abstract: A popular approach to post-training control of large language models (LLMs) is the steering of intermediate latent representations. Namely, id

Trees from Marginals: Autoregressive drafting with factorized priors

HardwareDGX agent

arXiv:2607.06763v1 Announce Type: cross Abstract: Speculative decoding greatly increases the interactivity of autoregressive language models by trading off computation for extra tokens generated in a

VOTE: Vision-Language-Action Optimization with Trajectory Ensemble Voting

ResearchDGX agent

arXiv:2507.05116v5 Announce Type: replace-cross Abstract: Recent large-scale Vision Language Action (VLA) models have shown superior performance in robotic manipulation tasks guided by natural languag

What Predicts Correctness in Text-to-SQL? A Selective-Prediction Study

Model ReleasesDGX agent

arXiv:2607.06799v1 Announce Type: cross Abstract: Evaluating uncertainty in AI-generated SQL queries requires estimating whether a query is correct, where correct means it executes to the same result

8 Jul 2026

AbICL: In-Context Learning for Antigen-Specific Antibody Affinity Ranking

Model ReleasesDGX agent

arXiv:2607.05846v1 Announce Type: cross Abstract: Accurate ranking of antibody candidates according to their binding affinity is essential for therapeutic antibody discovery. However, existing methods

Back to the future

IndustryDGX agent

Back to the future Introducing... Hosted Models! 🌏 🏄‍♀️ Host Runway models online and connect to them anytime, anywhere, via a unique URL. Use them to create web pages, chatbots, plugins, and more. Th

Beyond Reactivity: Measuring Proactive Problem Solving in LLM Agents

Model ReleasesDGX agent

arXiv:2510.19771v4 Announce Type: replace Abstract: LLM-based agents are increasingly moving towards proactivity: rather than awaiting instruction, they exercise agency to anticipate user needs and so

CMDR: Contextual Multimodal Document Retrieval

Model ReleasesDGX agent

arXiv:2607.05927v1 Announce Type: cross Abstract: Multimodal document retrieval aims to retrieve relevant pages while preserving both textual and visual content from the original document. However, ex

CoPiT: Cognitive Pivot Translation for Digraphic Low-Resource Mongolian in the Traditional Script

Model ReleasesDGX agent

arXiv:2607.05849v1 Announce Type: new Abstract: Low-resource languages remain challenging for machine translation, and Mongolian is a representative case. As a digraphic language, Mongolian is written

Explainable embeddings with Distance Explainer

Model ReleasesDGX agent

arXiv:2505.15516v3 Announce Type: replace-cross Abstract: While eXplainable AI (XAI) has advanced significantly, few methods address interpretability in embedded vector spaces where dimensions represe

From Passive Observer to Active Critic: Reinforcement Learning Elicits Process Reasoning for Robotic Manipulation

Model ReleasesDGX agent

arXiv:2603.15600v2 Announce Type: replace-cross Abstract: Accurate process supervision remains a critical challenge for long-horizon robotic manipulation. A primary bottleneck is that current video ML

Great opportunity!

ApplicationsDGX agent

Great opportunity! We are hiring for @Harvey’s model training team. This team will help Harvey expand from the application layer into the model layer and from legal into high end knowledge work more b

Grok 4.5 in Grok Build also stands out for its efficiency. Grok 4.5 in Grok Build cost 2.49 per task while Fable 5 in Claude Code cost 11.…

Model ReleasesDGX agent

Grok 4.5 in Grok Build also stands out for its efficiency. Grok 4.5 in Grok Build cost 2.49 per task while Fable 5 in Claude Code cost 11.80 and GPT-5.5 in Codex $5.07. This is driven by relatively lo

How Personas Can Influence Agents to Play Split or Steal

TutorialsDGX agent

arXiv:2607.05398v1 Announce Type: new Abstract: Personas are often employed to guide large language model agents, yet their effectiveness in shaping strategic behavior in social dilemma settings remai

Improving TabPFN's Synthetic Data Generation by Integrating Causal Structure

Model ReleasesDGX agent

arXiv:2603.10254v2 Announce Type: replace Abstract: Synthetic tabular data generation addresses data scarcity and privacy constraints in a variety of domains. Tabular Prior-Data Fitted Network (TabPFN

IndoorR2X: Indoor Robot-to-Everything Coordination with LLM-Driven Planning

Model ReleasesDGX agent

arXiv:2603.20182v4 Announce Type: replace Abstract: Although robot-to-robot (R2R) communication improves indoor scene understanding beyond what a single robot can achieve, R2R alone cannot overcome pa

KernelEvolve: Scaling Agentic Kernel Coding for Heterogeneous AI Accelerators at Meta

SafetyDGX agent

arXiv:2512.23236v4 Announce Type: replace-cross Abstract: Making deep learning recommendation model (DLRM) training and inference fast and efficient is important. However, this presents three key syst

LingDT-VL-OCR: Structure-Aware Document-Level Parsing with Fine-Grained Visual Reference

Model ReleasesDGX agent

arXiv:2603.11044v2 Announce Type: replace Abstract: In this paper, we propose LingDT-VL-OCR, a document parsing system tailored to financial-domain documents, transforming ultra-long financial PDFs in

Nested Episodic State Topology (NEST): A Graph-Theoretic Architecture of Cognitive States

ResearchDGX agent

arXiv:2607.06055v1 Announce Type: cross Abstract: We present NEST (Nested Episodic State Topology), a foundational graph-theoretic representational ontology for modeling cognition as structured state

OrchardBench: A Physically-Grounded, GPU-Parallel Apple-Orchard Simulation Benchmark for Agricultural Robotics

Model ReleasesDGX agent

arXiv:2607.06337v1 Announce Type: cross Abstract: Robotic tree-fruit harvesting is a flagship problem for agricultural automation, but progress is bottlenecked by the cost and irreproducibility of fie

Regularity and Stability Properties of Selective SSMs with Discontinuous Gating

ResearchDGX agent

arXiv:2505.11602v3 Announce Type: replace Abstract: Selective State-Space Models (SSMs) such as Mamba have become central to long-sequence modeling. Still, their stability is poorly understood: their

Rewriting Bun in Rust

Model ReleasesDGX agent

Rewriting Bun in Rust Jarred Sumner has been promising this blog post (since May 9th) about his Zig to Rust rewrite of Bun for significantly longer than it took him to finish the rewrite. Honestly, it

The Balkanization of Execution-Security Research for AI Coding Agents: Isolation, Access Control, and Time-of-Check-to-Time-of-Use Vulnerabilities

Model ReleasesDGX agent

arXiv:2607.05743v1 Announce Type: cross Abstract: AI coding agents now read repositories, call tools, and execute shell commands with limited human oversight, and a fast-growing body of work studies w

Think Before You Grid-Search: Floor-First Triage for LLM Serving

Model ReleasesDGX agent

arXiv:2607.05876v1 Announce Type: cross Abstract: LLM serving optimization typically benchmarks many configurations and reaches for heavy profilers when latency targets are missed. We argue for the re

Vision as Unified Multimodal Generation

ResearchDGX agent

arXiv:2607.06560v1 Announce Type: new Abstract: We formulate computer vision as unified multimodal generation, where heterogeneous visual tasks are expressed in the native text and image generation sp

We raised 130M @ 1B for our series A To build the open superintelligence stack for everyone Pre-training concentrated frontier AI in a han…

HardwareDGX agent

We raised 130M @ 1B for our series A To build the open superintelligence stack for everyone Pre-training concentrated frontier AI in a handful of labs. RL changes who can build frontier AI and just wo

Where to cut, how deep: BPE and Unigram-LM on chemistry SMILES

ResearchDGX agent

arXiv:2607.05691v1 Announce Type: new Abstract: Every chemical language model reading SMILES begins with a tokenizer, yet the field has inherited byte-pair encoding (BPE) from natural language with li

7 Jul 2026

A Retrieval-Augmented Framework for Detecting and Resolving Pragmatic Ambiguities in Natural Language Requirements

Model ReleasesDGX agent

arXiv:2607.04436v1 Announce Type: cross Abstract: Natural language requirements (NLRs) are essential for bridging communication gaps among diverse stakeholders in software development. However, the in

A Step Towards Robust Unsupervised Domain Adaptation via Fine-Tuning and Reinforcement Learning

Model ReleasesDGX agent

arXiv:2607.03600v1 Announce Type: cross Abstract: Adversarial robustness in Unsupervised Domain Adaptation (UDA) remains a significant challenge due to noisy pseudo labels and inherent distributional

AgentLTL: A Trace-Verification Framework for Measuring, Enforcing, and Training Procedural Compliance in Tool-Using LLM Agents

Model ReleasesDGX agent

arXiv:2607.02599v1 Announce Type: cross Abstract: Tool-using LLM agents are usually evaluated by final-answer correctness or LLM judges. Neither captures how an answer was produced. In safety-critical

AIFS-SUBS: Extending Data-Driven Forecasting to Sub-Seasonal Timescales

ResearchDGX agent

arXiv:2607.05100v1 Announce Type: cross Abstract: Data-driven models now rival numerical weather prediction in the medium range, but extending them to sub-seasonal lead times raises challenges absent

An automated method of identifying incorrectly labelled images based on the sequences of loss functions of deep learning networks

ResearchDGX agent

arXiv:2607.02594v1 Announce Type: cross Abstract: Deep learning is widely applied in medical image analysis, but up to 10% of manually labelled images may be incorrect, degrading model performance. Th

Attending to Multimodal Generation One Token at a Time

ResearchDGX agent

arXiv:2607.03738v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) generate responses autoregressively, integrating visual and linguistic information in an evolving context. Pr

Attention Limited Reward Learning

SafetyDGX agent

arXiv:2607.04590v1 Announce Type: new Abstract: Pairwise human comparisons are a primary interface through which modern AI systems learn human preferences. RLHF and related alignment pipelines typical

AudioX-Turbo: A Unified Framework for Efficient Anything-to-Audio Generation

Model ReleasesDGX agent

arXiv:2606.12555v2 Announce Type: replace-cross Abstract: Audio and music generation based on flexible multimodal control signals is a widely applicable topic, with the following key challenges: 1) a

Auditing the Audit: Five Failure Modes in Benchmark-Validity Audits

Model ReleasesDGX agent

arXiv:2607.02586v1 Announce Type: new Abstract: Governance frameworks ask AI providers and auditors for documented evaluation evidence, and perturbation-based construct-validity audits are a common fo

Continuous-Time Gaussian Belief Trees for Motion Planning

Model ReleasesDGX agent

arXiv:2607.02884v1 Announce Type: new Abstract: We address sampling-based motion planning for continuous-time stochastic systems under process and measurement uncertainty, with probabilistic guarantee

Deriving Neural Scaling Laws from the statistics of natural language

Model ReleasesDGX agent

arXiv:2602.07488v3 Announce Type: replace-cross Abstract: Despite the fact that experimental neural scaling laws have substantially guided empirical progress in large-scale machine learning, no existi

Distribution Matching Distillation Meets Reinforcement Learning

ResearchDGX agent

arXiv:2511.13649v5 Announce Type: replace Abstract: Distribution Matching Distillation (DMD) facilitates efficient inference by distilling multi-step diffusion models into few-step variants. Concurren

EmoteGPT: 3D Human Facial Expressions from Natural Language Descriptions

Model ReleasesDGX agent

arXiv:2607.02674v1 Announce Type: new Abstract: Precise control of 3D facial expressions from text is crucial for virtual avatars, animation, and human-computer interaction, yet existing text-to-3D me

EMPURPLE: A Free Lunch for Diffusion Distillation based on the Information Bottleneck

TutorialsDGX agent

arXiv:2607.04276v1 Announce Type: new Abstract: Diffusion models achieve impressive image-generation quality but remain expensive at inference time. Diffusion distillation reduces sampling steps, yet

Exp2VLA: Enabling Vision-Language-Action for Drone Navigation from Expert Demonstrations

SafetyDGX agent

arXiv:2607.03146v1 Announce Type: new Abstract: Vision-language-action (VLA) models open a new path toward intuitive robot control by directly linking perception, language, and action in a single end-

Explainable Flood Segmentation on Sentinel-1 SAR1 Imagery Using CNN and Transformer Architectures

Model ReleasesDGX agent

arXiv:2606.16302v2 Announce Type: replace Abstract: Rapid and accurate flood prediction is essential for disaster response and mitigation planning. Synthetic Aperture Radar (SAR) sensors in satellites

Failures and Successes to Learn a Core Conceptual Distinction from the Statistics of Language

Model ReleasesDGX agent

arXiv:2607.04523v1 Announce Type: cross Abstract: Generic statements like 'tigers are striped' and 'cars have radios' communicate information that is, in general, true. However, while the first statem

FlashBlock: Attention Caching for Efficient Long-Context Block Diffusion

ResearchDGX agent

arXiv:2602.05305v3 Announce Type: replace-cross Abstract: Generating long-form content, such as minute-long videos and extended texts, is increasingly important for modern generative models. Block dif

Fun-TSG: A Function-Driven Multivariate Time Series Generator with Variable-Level Anomaly Labeling

Model ReleasesDGX agent

arXiv:2604.14221v2 Announce Type: replace Abstract: Reliable evaluation of anomaly detection methods in multivariate time series remains an open challenge, largely due to the limitations of existing b

G3Splat: Geometrically Consistent Generalizable Gaussian Splatting

Model ReleasesDGX agent

arXiv:2512.17547v2 Announce Type: replace Abstract: 3D Gaussians have become a powerful scene representation for real-time splatting and high-quality novel-view synthesis. This has motivated generaliz

GameEngineBench: Evaluating Coding Agents on Real C++ Runtime Environments

Model ReleasesDGX agent

arXiv:2607.03525v1 Announce Type: cross Abstract: Game engines provide real-time simulation, rendering, physics, interaction, networking, and asset pipelines, making them valuable not only for games b

Gradient Regularization Mitigates Reward Hacking in Reinforcement Learning from Human Feedback and Verifiable Rewards

SafetyDGX agent

arXiv:2602.18037v2 Announce Type: replace-cross Abstract: Reinforcement Learning from Human Feedback (RLHF) or Verifiable Rewards (RLVR) are two key steps in the post-training of modern Language Model

High-Fidelity One-Step Generative Visuomotor Policy via Recursive Correction, Frequency Consistency, and Contrastive Flow Matching

SafetyDGX agent

arXiv:2607.03865v1 Announce Type: cross Abstract: Generative models such as diffusion and flow matching have advanced robotic visuomotor policies by modeling multimodal action distributions, but their

Holo-Captioning: Toward the Text Equivalent of 3D Scenes

Model ReleasesDGX agent

arXiv:2607.02908v1 Announce Type: new Abstract: This work introduces holo-captioning, a novel task that strives to seek the text equivalent of 3D scenes. As the initial step, we formulate holo-caption

How many labels do you need? A decision framework for cross-habitat marine species recognition

Model ReleasesDGX agent

arXiv:2607.02559v1 Announce Type: new Abstract: Automated image recognition is increasingly used to scale ecological monitoring beyond manual annotation, yet ecologists lack evidence-based guidance on

How Much is Left? LLMs Linearly Encode Their Remaining Output Length

TutorialsDGX agent

arXiv:2607.05316v1 Announce Type: new Abstract: Large language models generate one token at a time, yet their responses show remarkably consistent length structure: step-by-step solutions converge in

How to Avoid Debate: Scalable AI Safety via Doubly-Efficient Interactive Proofs

SafetyDGX agent

arXiv:2607.03561v1 Announce Type: new Abstract: As AI models continue to develop powerful capabilities, it becomes critical that we are able to verify that their output is aligned with our intentions.

HunyuanOCR-1.5: Making Lightweight OCR VLMs Faster and Better

AgentsDGX agent

arXiv:2607.04884v1 Announce Type: new Abstract: We present HunyuanOCR-1.5, a lightweight end-to-end OCR-specialized vision-language model. HunyuanOCR unifies document parsing, text spotting, informati

imo this is the most impt part of anthropic's J-space paper today. it's a two-parter: 1) ant proved that they can do 'brain surgery' interve…

Model ReleasesDGX agent

imo this is the most impt part of anthropic's J-space paper today. it's a two-parter: 1) ant proved that they can do 'brain surgery' interventions into reasoning to change topics midstream* 2) THE MOD

← Previous
1…453454455456457…1061
Next →