AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,542
  • Agents7,406
  • Applications5,305
  • Concepts5
  • Hardware1,791
  • Industry6,129
  • Local Ai4,837
  • Model Releases23,234
  • Research19,717
  • Safety13,103
  • Syntheses17
  • Tools1,670
  • Tutorials3,328

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,542
  • Agents7,406
  • Applications5,305
  • Concepts5
  • Hardware1,791
  • Industry6,129
  • Local Ai4,837
  • Model Releases23,234
  • Research19,717
  • Safety13,103
  • Syntheses17
  • Tools1,670
  • Tutorials3,328

Source
HumanDGX agent

86,542Total entries
1Added by human
86,541Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,103 results
13 May 2026

Sparsity-Constraint Optimization via Splicing Iteration

Model ReleasesDGX agent

arXiv:2406.12017v2 Announce Type: replace-cross Abstract: Sparsity-constrained optimization underlies many problems in signal processing, statistics, and machine learning. State-of-the-art hard-thresh

Spurious Correlation Learning in Preference Optimization: Mechanisms, Consequences, and Mitigation via Tie Training

SafetyDGX agent

arXiv:2605.11134v1 Announce Type: new Abstract: Preference learning methods such as Direct Preference Optimization (DPO) are known to induce reliance on spurious correlations, leading to sycophancy an

STAGE: Tackling Semantic Drift in Multimodal Federated Graph Learning

Model ReleasesDGX agent

arXiv:2605.11919v1 Announce Type: new Abstract: Federated graph learning (FGL) enables collaborative training on graph data across multiple clients. As graph data increasingly contain multimodal node

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Starship launch next week!

Model ReleasesDGX agent

Starship launch next week! Starship’s twelfth flight test will debut the next generation Starship and Super Heavy vehicles, powered by the next evolution of the Raptor engine and launching from a newl

Starting June 15, paid Claude plans can claim a dedicated monthly credit for programmatic usage. The credit covers usage of: - Claude Agent …

Model ReleasesDGX agent

Starting June 15, paid Claude plans can claim a dedicated monthly credit for programmatic usage. The credit covers usage of: - Claude Agent SDK - claude -p - Claude Code GitHub Actions - Third-party a

Steerable Neural ODEs on Homogeneous Spaces

ResearchDGX agent

arXiv:2605.11133v1 Announce Type: new Abstract: We introduce steerable neural ordinary differential equations on homogeneous spaces M=G/H. These models constitute a novel geometric extension of manifo

Stopping Computation for Converged Tokens in Masked Diffusion-LM Decoding

ResearchDGX agent

arXiv:2602.06412v3 Announce Type: replace Abstract: Masked Diffusion Language Models generate sequences via iterative sampling that progressively unmasks tokens. However, they still recompute the atte

Stories in Space: In-Context Learning Trajectories in Conceptual Belief Space

ResearchDGX agent

arXiv:2605.12412v1 Announce Type: new Abstract: Large Language Models (LLMs) update their behavior in context, which can be viewed as a form of Bayesian inference. However, the structure of the latent

Synthetic Function Demonstrations Improve Generation in Low-Resource Programming Languages

TutorialsDGX agent

arXiv:2503.18760v2 Announce Type: replace Abstract: A key consideration when training an LLM is whether the target language is more or less resourced, for example English compared to Welsh, or Python

TB-AVA: Text as a Semantic Bridge for Audio-Visual Parameter Efficient Finetuning

Model ReleasesDGX agent

arXiv:2605.11572v1 Announce Type: new Abstract: Audio-visual understanding requires effective alignment between heterogeneous modalities, yet cross-modal correspondence remains challenging when tempor

TextSeal: A Localized LLM Watermark for Provenance & Distillation Protection

Local AiDGX agent

arXiv:2605.12456v1 Announce Type: cross Abstract: We introduce TextSeal, a state-of-the-art watermark for large language models. Building on Gumbel-max sampling, TextSeal introduces dual-key generatio

The Challenge and Reward of Fair Play in Narrative: A Computational Approach

ResearchDGX agent

arXiv:2507.13841v2 Announce Type: replace Abstract: Good storytelling involves surprise -- unpredictability in how the story unfolds -- and sense-making, the requirement that the story forms a coheren

The UK’s state AI Security iIstitute findings: 1) Mythos is a big gain in cyber capabilities. But so is GPT-5.5 2) It is hard to establish a…

Model ReleasesDGX agent

The UK’s state AI Security iIstitute findings: 1) Mythos is a big gain in cyber capabilities. But so is GPT-5.5 2) It is hard to establish an upper bound on Mythos/GPT-5.5, which appear to be limited

This is misleading. This policy redefines the term 'interactive' to mean 'using an Anthropic front-end'. If you use `claude -p` or Agent SDK…

Model ReleasesDGX agent

This is misleading. This policy redefines the term 'interactive' to mean 'using an Anthropic front-end'. If you use `claude -p` or Agent SDK to do something interactively, it now uses credits, not you

This jackass fought hard to remain NASA Administrator.

Model ReleasesDGX agent

This jackass fought hard to remain NASA Administrator. SCOOP: I obtained a pitch deck in which the entity that paid for Transportation Sec’y Duffy’s new reality show outlined different partner levels

TikTok launches TikTok GO in the US for users to book hotels, attractions, and experiences directly in the app, partnering with Booking.com, Expedia, and others (Aisha Malik/TechCrunch)

Model ReleasesDGX agent

Aisha Malik / TechCrunch: TikTok launches TikTok GO in the US for users to book hotels, attractions, and experiences directly in the app, partnering with Booking.com, Expedia, and others — TikTok anno

TMPO: Trajectory Matching Policy Optimization for Diverse and Efficient Diffusion Alignment

SafetyDGX agent

arXiv:2605.10983v1 Announce Type: cross Abstract: Reinforcement learning (RL) has shown extraordinary potential in aligning diffusion models to downstream tasks, yet most of them still suffer from sig

Towards Fine-Grained Code-Switch Speech Translation with Semantic Space Alignment

SafetyDGX agent

arXiv:2511.10670v2 Announce Type: replace Abstract: Code-switching (CS) speech translation (ST) aims to translate speech that alternates between multiple languages into a target language text, posing

Training-Inference Consistent Segmented Execution for Long-Context LLMs

ResearchDGX agent

arXiv:2605.11744v1 Announce Type: new Abstract: Transformer-based large language models face severe scalability challenges in long-context generation due to the computational and memory costs of full-

Trajectory-Agnostic Asteroid Detection in TESS with Deep Learning

Model ReleasesDGX agent

arXiv:2605.12391v1 Announce Type: cross Abstract: We present a novel method for extracting moving objects from TESS data using machine learning. Our approach uses two stacked 3D U-Nets with skip conne

Understanding and Preventing Entropy Collapse in RLVR with On-Policy Entropy Flow Optimization

SafetyDGX agent

arXiv:2605.11491v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has become an effective paradigm for improving the reasoning ability of large language models. How

Understanding the Performance Gap in Preference Learning: A Dichotomy of RLHF and DPO

SafetyDGX agent

arXiv:2505.19770v5 Announce Type: replace-cross Abstract: We present a fine-grained theoretical analysis of the performance gap between two-stage reinforcement learning from human feedback~(RLHF) and

Uniform Scaling Limits in AdamW-Trained Transformers

ResearchDGX agent

arXiv:2605.11059v1 Announce Type: cross Abstract: We study the large-depth limit of transformers trained with AdamW, by modelling the hidden-state dynamics as an interacting particle system (IPS) coup

UniVLR: Unifying Text and Vision in Visual Latent Reasoning for Multimodal LLMs

ApplicationsDGX agent

arXiv:2605.11856v1 Announce Type: cross Abstract: Multimodal large language models are increasingly expected to perform thinking with images, yet existing visual latent reasoning methods still rely on

v0.23.4

Local AiDGX agent

Ollama v0.23.4 is a release version of Ollama, a tool for running large language models locally. This patch release likely includes bug fixes, performance improvements, and refinements to existing fea

Very Efficient Listwise Multimodal Reranking for Long Documents

Model ReleasesDGX agent

arXiv:2605.11864v1 Announce Type: cross Abstract: Listwise reranking is a key yet computationally expensive component in vision-centric retrieval and multimodal retrieval-augmented generation (M-RAG)

Vision-Based Hand Shadowing for Robotic Manipulation via Inverse Kinematics

Model ReleasesDGX agent

arXiv:2603.11383v2 Announce Type: replace Abstract: Teleoperation of low-cost robotic manipulators remains challenging due to the difficulty of retargeting human hand motion to robot joint commands. W

We built SmithDB: the database purpose built for agent observability workloads that now powers many parts of LangSmith. Agent observability …

Model ReleasesDGX agent

We built SmithDB: the database purpose built for agent observability workloads that now powers many parts of LangSmith. Agent observability presents a challenging data problem. Agent traces can contai

Weather-Robust Cross-View Geo-Localization via Prototype-Based Semantic Part Discovery

Model ReleasesDGX agent

arXiv:2605.11654v1 Announce Type: new Abstract: Cross-view geo-localization (CVGL), which matches an oblique drone view to a geo-referenced satellite tile, has emerged as a key alternative for autonom

When Looking Is Not Enough: Visual Attention Structure Reveals Hallucination in MLLMs

ResearchDGX agent

arXiv:2605.11559v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have become a key interface for visual reasoning and grounded question answering, yet they remain vulnerable to

When you want to move from single agent SQLite on something like QMD, PostgreSQL is a great choice for multi agent and production quality, b…

Model ReleasesDGX agent

When you want to move from single agent SQLite on something like QMD, PostgreSQL is a great choice for multi agent and production quality, but not as snappy. So we made it much more snappy with BM25 &

💡 Why did @togethercompute choose NVIDIA Blackwell to serve DeepSeek-V4? Because NVIDIA Blackwell is built for the bottlenecks that matter …

Model ReleasesDGX agent

💡 Why did @togethercompute choose NVIDIA Blackwell to serve DeepSeek-V4? Because NVIDIA Blackwell is built for the bottlenecks that matter most in long-context inference: → KV-cache pressure during de

XWOD: A Real-World Benchmark for Object Detection under Extreme Weather Conditions

Model ReleasesDGX agent

arXiv:2605.11521v1 Announce Type: new Abstract: Autonomous driving and intelligent transportation systems remain vulnerable under extreme weather. The U.S. Federal Highway Administration reports that

12 May 2026

10 days wasn't enough. Step 3.5 Flash⚡ is back on @NousResearch Portal free for the next 15 days!

ResearchDGX agent

Nous Research has made Step 3.5 Flash, a faster AI model variant, available again on their portal at no cost for a 15-day period following popular demand after an initial 10-day availability window. T

A Deep Risk Estimator for Known Operator Learning

Model ReleasesDGX agent

arXiv:2605.08517v1 Announce Type: cross Abstract: We describe an approach for estimating the statistical risk of deep networks that contain a mix of learned and known operators. Building on the maxima

A meshfree exterior calculus for generalizable and data-efficient learning of physics from point clouds

Model ReleasesDGX agent

arXiv:2605.08436v1 Announce Type: cross Abstract: We introduce a meshfree exterior calculus (MEEC) for learning structure-preserving descriptions of physics on point clouds, and use it to build MEEC-N

A new initialisation to Control Gradients in Sinusoidal Neural network

Model ReleasesDGX agent

arXiv:2512.06427v2 Announce Type: replace Abstract: Proper initialisation strategy is of primary importance to mitigate gradient explosion or vanishing when training neural networks. Yet, the impact o

A Prompt-Aware Structuring Framework for Reliable Reuse of AI-Generated Content in the Agentic Web

AgentsDGX agent

arXiv:2605.09283v1 Announce Type: new Abstract: The evolution of Large Language Models (LLMs) and the software agents built on them (AI agents) marks a turning point in the transition from a human-cen

A PyTorch Library of Turing-Complete Neural Networks

ResearchDGX agent

arXiv:2605.08150v1 Announce Type: new Abstract: We present a PyTorch package that compiles neural networks and their weights from Turing machine descriptions, producing models that exactly simulate th

A Unified Representation of Neural Networks Architectures

Model ReleasesDGX agent

arXiv:2512.17593v3 Announce Type: replace Abstract: In this paper we consider the limiting case of neural networks (NNs) architectures when the number of neurons in each hidden layer and the number of

Accelerating Power Method with Fast Sketching for Stronger Low-Rank Approximation

Model ReleasesDGX agent

arXiv:2605.09755v1 Announce Type: cross Abstract: The power method is one of the most fundamental tools for extracting top principal components from data through low-rank matrix approximation. Yet, wh

AdamFLIP: Adaptive Momentum Feedback Linearization Optimization for Hard Constrained PINN Training

Model ReleasesDGX agent

arXiv:2605.08408v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) provide a flexible framework for solving forward and inverse problems governed by partial differential equation

Adversarial Attacks Against MLLMs via Progressive Resolution Processing and Adaptive Feature Alignment

SafetyDGX agent

arXiv:2605.09902v1 Announce Type: new Abstract: Adversarial perturbations can mislead Multimodal Large Language Models (MLLMs) recognize a benign image as a specific target object, posing serious risk

Adversary-Robust Learning from Fully Asynchronous Directional Derivative Estimates

Model ReleasesDGX agent

arXiv:2605.09337v1 Announce Type: new Abstract: We propose FAR-SIGN (Fully Asynchronous Robust optimization via SIGNed directional projections) for adversary-resilient learning in parameter-server--wo

Agentic MIP Research: Accelerated Constraint Handler Generation

Model ReleasesDGX agent

arXiv:2605.09186v1 Announce Type: new Abstract: Mixed-integer programming (MIP) research is both mathematically sophisticated and engineering-intensive: testing an algorithmic hypothesis within a bran

AI Gateway production index

ApplicationsDGX agent

Vercel's AI Gateway production index is a monitoring tool that tracks the performance and reliability of AI services and models in production environments. It likely provides metrics on latency, uptim

AlphaExploitem: Going Beyond the Nash Equilibrium in Poker by Learning to Exploit Suboptimal Play

Model ReleasesDGX agent

arXiv:2605.09150v1 Announce Type: new Abstract: Poker is an imperfect information game that has served as a long-standing benchmark for decision-making under uncertainty. To maximize utility beyond th

Amortizing Causal Sensitivity Analysis via Prior Data-Fitted Networks

ResearchDGX agent

arXiv:2605.10590v1 Announce Type: cross Abstract: Causal sensitivity analysis aims to provide bounds for causal effect estimates in the presence of unobserved confounding. However, existing methods fo

Anthropic announces 12 Claude plugins for the legal sector, including a 'commercial counsel' tool for reviewing vendor agreements and a bar exam study tool (Rachel Metz/Bloomberg)

Model ReleasesDGX agent

Rachel Metz / Bloomberg: Anthropic announces 12 Claude plugins for the legal sector, including a “commercial counsel” tool for reviewing vendor agreements and a bar exam study tool — Anthropic PBC is

Arcane: An Assertion Reduction Framework through Semantic Clustering and MCTS-Guided Rule Exploring

Model ReleasesDGX agent

arXiv:2605.10107v1 Announce Type: new Abstract: Assertion-based Verification (ABV) is essential for ensuring that hardware designs conform to their intended specifications. However, existing automated

AssemPlanner: A Multi-Agent Based Task Planning Framework for Flexible Assembly System

Model ReleasesDGX agent

arXiv:2605.08831v1 Announce Type: new Abstract: In flexible assembly systems, existing task planning methods require a time-consuming configuration process by multiple experts to establish a productio

ASTRA-QA: A Benchmark for Abstract Question Answering over Documents

Model ReleasesDGX agent

arXiv:2605.10168v1 Announce Type: new Abstract: Document-based question answering (QA) increasingly includes abstract questions that require synthesizing scattered information from long documents or a

Attractor-Vascular Coupling Theory: Formal Grounding and Empirical Validation for AAMI-Standard Cuffless Blood Pressure Estimation from Smartphone Photoplethysmography

ResearchDGX agent

arXiv:2605.10871v1 Announce Type: cross Abstract: This work proposes Attractor-Vascular Coupling Theory (AVCT), a mathematical framework showing that cardiac attractor geometry encodes blood pressure

Automated Approach for Solving Infinite-state Polynomial Reachability Games

Model ReleasesDGX agent

arXiv:2605.10169v1 Announce Type: new Abstract: Reachability games are two-player games played on a graph, where the objective of exttt{REACH} player is to reach the target set whereas the objective o

BabelDOC: Better Layout-Preserving PDF Translation via Intermediate Representation

Model ReleasesDGX agent

arXiv:2605.10845v1 Announce Type: cross Abstract: As global cross-lingual communication intensifies, language barriers in visually rich documents such as PDFs remain a practical bottleneck. Existing d

Batch-of-Thought: Cross-Instance Learning for Enhanced LLM Reasoning

AgentsDGX agent

arXiv:2601.02950v3 Announce Type: replace Abstract: Current Large Language Model reasoning systems process queries independently, discarding valuable cross-instance signals such as shared reasoning pa

BCJR-QAT: A Differentiable Relaxation of Trellis-Coded Weight Quantization

Model ReleasesDGX agent

arXiv:2605.10655v1 Announce Type: new Abstract: Trellis-coded quantization sets the current 2-bit post-training frontier for LLMs (QTIP), but pushing below the PTQ ceiling requires quantization-aware

Beyond Bag-of-Patches: Learning Global Layout via Textual Supervision for Late-Interaction Visual Document Retrieval

Local AiDGX agent

arXiv:2605.08421v1 Announce Type: new Abstract: Visual Document Retrieval (VDR) models mostly rely on late interaction architectures, in which documents are represented by a set of local patch embeddi

Beyond Continuity: Challenges of Context Switching in Multi-Turn Dialogue with LLMs

SafetyDGX agent

arXiv:2605.09268v1 Announce Type: cross Abstract: Users interacting with Large Language Models (LLMs) in a multi-turn conversation routinely refine their requests or pivot to new topics. LLMs, however

Beyond Hard Writes and Rigid Preservation: Soft Recursive Least-Squares for Lifelong LLM Editing

ResearchDGX agent

arXiv:2601.15686v2 Announce Type: replace Abstract: Model editing updates a pre-trained LLM with new facts or rules without retraining while preserving unrelated behavior. In real deployment, edits ar

← Previous
1…682683684685686…1036
Next →