AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,573
  • Agents7,495
  • Applications5,364
  • Concepts5
  • Hardware1,813
  • Industry6,149
  • Local Ai4,892
  • Model Releases23,569
  • Research19,967
  • Safety13,263
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,573
  • Agents7,495
  • Applications5,364
  • Concepts5
  • Hardware1,813
  • Industry6,149
  • Local Ai4,892
  • Model Releases23,569
  • Research19,967
  • Safety13,263
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
AllBlog
87,573Total entries
1Added by human
87,572Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,952 results
Research

Steerable Neural ODEs on Homogeneous Spaces

DGX agent

arXiv:2605.11133v1 Announce Type: new Abstract: We introduce steerable neural ordinary differential equations on homogeneous spaces M=G/H. These models constitute a novel geometric extension of manifo

researcharxiv-cs-lg
13 May 2026
Research

Stopping Computation for Converged Tokens in Masked Diffusion-LM Decoding

DGX agent
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

arXiv:2602.06412v3 Announce Type: replace Abstract: Masked Diffusion Language Models generate sequences via iterative sampling that progressively unmasks tokens. However, they still recompute the atte

researcharxiv-cs-cl
13 May 2026
Research

Stories in Space: In-Context Learning Trajectories in Conceptual Belief Space

DGX agent

arXiv:2605.12412v1 Announce Type: new Abstract: Large Language Models (LLMs) update their behavior in context, which can be viewed as a form of Bayesian inference. However, the structure of the latent

researcharxiv-cs-cl
13 May 2026
Tutorials

Synthetic Function Demonstrations Improve Generation in Low-Resource Programming Languages

DGX agent

arXiv:2503.18760v2 Announce Type: replace Abstract: A key consideration when training an LLM is whether the target language is more or less resourced, for example English compared to Welsh, or Python

tutorialsarxiv-cs-cl
13 May 2026
Model Releases

TB-AVA: Text as a Semantic Bridge for Audio-Visual Parameter Efficient Finetuning

DGX agent

arXiv:2605.11572v1 Announce Type: new Abstract: Audio-visual understanding requires effective alignment between heterogeneous modalities, yet cross-modal correspondence remains challenging when tempor

model-releasesarxiv-cs-cv
13 May 2026
Local Ai

TextSeal: A Localized LLM Watermark for Provenance & Distillation Protection

DGX agent

arXiv:2605.12456v1 Announce Type: cross Abstract: We introduce TextSeal, a state-of-the-art watermark for large language models. Building on Gumbel-max sampling, TextSeal introduces dual-key generatio

local-aiarxiv-cs-cl
13 May 2026
Research

The Challenge and Reward of Fair Play in Narrative: A Computational Approach

DGX agent

arXiv:2507.13841v2 Announce Type: replace Abstract: Good storytelling involves surprise -- unpredictability in how the story unfolds -- and sense-making, the requirement that the story forms a coheren

researcharxiv-cs-cl
13 May 2026
Model Releases

The UK’s state AI Security iIstitute findings: 1) Mythos is a big gain in cyber capabilities. But so is GPT-5.5 2) It is hard to establish a…

DGX agent

The UK’s state AI Security iIstitute findings: 1) Mythos is a big gain in cyber capabilities. But so is GPT-5.5 2) It is hard to establish an upper bound on Mythos/GPT-5.5, which appear to be limited

model-releasesethan-mollick--x
13 May 2026
Model Releases

This is misleading. This policy redefines the term 'interactive' to mean 'using an Anthropic front-end'. If you use `claude -p` or Agent SDK…

DGX agent

This is misleading. This policy redefines the term 'interactive' to mean 'using an Anthropic front-end'. If you use `claude -p` or Agent SDK to do something interactively, it now uses credits, not you

model-releasesjeremy-howard--x
13 May 2026
Model Releases

This jackass fought hard to remain NASA Administrator.

DGX agent

This jackass fought hard to remain NASA Administrator. SCOOP: I obtained a pitch deck in which the entity that paid for Transportation Sec’y Duffy’s new reality show outlined different partner levels

model-releasesanthropic--x
13 May 2026
Model Releases

TikTok launches TikTok GO in the US for users to book hotels, attractions, and experiences directly in the app, partnering with Booking.com, Expedia, and others (Aisha Malik/TechCrunch)

DGX agent

Aisha Malik / TechCrunch: TikTok launches TikTok GO in the US for users to book hotels, attractions, and experiences directly in the app, partnering with Booking.com, Expedia, and others — TikTok anno

model-releasestechmeme
13 May 2026
Safety

TMPO: Trajectory Matching Policy Optimization for Diverse and Efficient Diffusion Alignment

DGX agent

arXiv:2605.10983v1 Announce Type: cross Abstract: Reinforcement learning (RL) has shown extraordinary potential in aligning diffusion models to downstream tasks, yet most of them still suffer from sig

safetyarxiv-cs-cv
13 May 2026
Safety

Towards Fine-Grained Code-Switch Speech Translation with Semantic Space Alignment

DGX agent

arXiv:2511.10670v2 Announce Type: replace Abstract: Code-switching (CS) speech translation (ST) aims to translate speech that alternates between multiple languages into a target language text, posing

safetyarxiv-cs-cl
13 May 2026
Research

Training-Inference Consistent Segmented Execution for Long-Context LLMs

DGX agent

arXiv:2605.11744v1 Announce Type: new Abstract: Transformer-based large language models face severe scalability challenges in long-context generation due to the computational and memory costs of full-

researcharxiv-cs-cl
13 May 2026
Model Releases

Trajectory-Agnostic Asteroid Detection in TESS with Deep Learning

DGX agent

arXiv:2605.12391v1 Announce Type: cross Abstract: We present a novel method for extracting moving objects from TESS data using machine learning. Our approach uses two stacked 3D U-Nets with skip conne

model-releasesarxiv-cs-lg
13 May 2026
Safety

Understanding and Preventing Entropy Collapse in RLVR with On-Policy Entropy Flow Optimization

DGX agent

arXiv:2605.11491v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has become an effective paradigm for improving the reasoning ability of large language models. How

safetyarxiv-cs-lg
13 May 2026
Safety

Understanding the Performance Gap in Preference Learning: A Dichotomy of RLHF and DPO

DGX agent

arXiv:2505.19770v5 Announce Type: replace-cross Abstract: We present a fine-grained theoretical analysis of the performance gap between two-stage reinforcement learning from human feedback~(RLHF) and

safetyarxiv-cs-cl
13 May 2026
Research

Uniform Scaling Limits in AdamW-Trained Transformers

DGX agent

arXiv:2605.11059v1 Announce Type: cross Abstract: We study the large-depth limit of transformers trained with AdamW, by modelling the hidden-state dynamics as an interacting particle system (IPS) coup

researcharxiv-cs-lg
13 May 2026
Applications

UniVLR: Unifying Text and Vision in Visual Latent Reasoning for Multimodal LLMs

DGX agent

arXiv:2605.11856v1 Announce Type: cross Abstract: Multimodal large language models are increasingly expected to perform thinking with images, yet existing visual latent reasoning methods still rely on

applicationsarxiv-cs-cl
13 May 2026
Local Ai

v0.23.4

DGX agent

Ollama v0.23.4 is a release version of Ollama, a tool for running large language models locally. This patch release likely includes bug fixes, performance improvements, and refinements to existing fea

local-aiollama-releases
13 May 2026
Model Releases

Very Efficient Listwise Multimodal Reranking for Long Documents

DGX agent

arXiv:2605.11864v1 Announce Type: cross Abstract: Listwise reranking is a key yet computationally expensive component in vision-centric retrieval and multimodal retrieval-augmented generation (M-RAG)

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Vision-Based Hand Shadowing for Robotic Manipulation via Inverse Kinematics

DGX agent

arXiv:2603.11383v2 Announce Type: replace Abstract: Teleoperation of low-cost robotic manipulators remains challenging due to the difficulty of retargeting human hand motion to robot joint commands. W

model-releasesarxiv-cs-ro
13 May 2026
Model Releases

We built SmithDB: the database purpose built for agent observability workloads that now powers many parts of LangSmith. Agent observability …

DGX agent

We built SmithDB: the database purpose built for agent observability workloads that now powers many parts of LangSmith. Agent observability presents a challenging data problem. Agent traces can contai

model-releasesharrison-chase--x
13 May 2026
Model Releases

Weather-Robust Cross-View Geo-Localization via Prototype-Based Semantic Part Discovery

DGX agent

arXiv:2605.11654v1 Announce Type: new Abstract: Cross-view geo-localization (CVGL), which matches an oblique drone view to a geo-referenced satellite tile, has emerged as a key alternative for autonom

model-releasesarxiv-cs-cv
13 May 2026
Research

When Looking Is Not Enough: Visual Attention Structure Reveals Hallucination in MLLMs

DGX agent

arXiv:2605.11559v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have become a key interface for visual reasoning and grounded question answering, yet they remain vulnerable to

researcharxiv-cs-cv
13 May 2026
Model Releases

When you want to move from single agent SQLite on something like QMD, PostgreSQL is a great choice for multi agent and production quality, b…

DGX agent

When you want to move from single agent SQLite on something like QMD, PostgreSQL is a great choice for multi agent and production quality, but not as snappy. So we made it much more snappy with BM25 &

model-releasesemad-mostaque--x
13 May 2026
Model Releases

💡 Why did @togethercompute choose NVIDIA Blackwell to serve DeepSeek-V4? Because NVIDIA Blackwell is built for the bottlenecks that matter …

DGX agent

💡 Why did @togethercompute choose NVIDIA Blackwell to serve DeepSeek-V4? Because NVIDIA Blackwell is built for the bottlenecks that matter most in long-context inference: → KV-cache pressure during de

model-releasestogether-ai--x
13 May 2026
Model Releases

XWOD: A Real-World Benchmark for Object Detection under Extreme Weather Conditions

DGX agent

arXiv:2605.11521v1 Announce Type: new Abstract: Autonomous driving and intelligent transportation systems remain vulnerable under extreme weather. The U.S. Federal Highway Administration reports that

model-releasesarxiv-cs-cv
13 May 2026
Research

10 days wasn't enough. Step 3.5 Flash⚡ is back on @NousResearch Portal free for the next 15 days!

DGX agent

Nous Research has made Step 3.5 Flash, a faster AI model variant, available again on their portal at no cost for a 15-day period following popular demand after an initial 10-day availability window. T

researchnous-research--x
12 May 2026
Model Releases

A Deep Risk Estimator for Known Operator Learning

DGX agent

arXiv:2605.08517v1 Announce Type: cross Abstract: We describe an approach for estimating the statistical risk of deep networks that contain a mix of learned and known operators. Building on the maxima

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

A meshfree exterior calculus for generalizable and data-efficient learning of physics from point clouds

DGX agent

arXiv:2605.08436v1 Announce Type: cross Abstract: We introduce a meshfree exterior calculus (MEEC) for learning structure-preserving descriptions of physics on point clouds, and use it to build MEEC-N

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

A new initialisation to Control Gradients in Sinusoidal Neural network

DGX agent

arXiv:2512.06427v2 Announce Type: replace Abstract: Proper initialisation strategy is of primary importance to mitigate gradient explosion or vanishing when training neural networks. Yet, the impact o

model-releasesarxiv-cs-lg
12 May 2026
Agents

A Prompt-Aware Structuring Framework for Reliable Reuse of AI-Generated Content in the Agentic Web

DGX agent

arXiv:2605.09283v1 Announce Type: new Abstract: The evolution of Large Language Models (LLMs) and the software agents built on them (AI agents) marks a turning point in the transition from a human-cen

agentsarxiv-cs-ai
12 May 2026
Research

A PyTorch Library of Turing-Complete Neural Networks

DGX agent

arXiv:2605.08150v1 Announce Type: new Abstract: We present a PyTorch package that compiles neural networks and their weights from Turing machine descriptions, producing models that exactly simulate th

researcharxiv-cs-lg
12 May 2026
Model Releases

A Unified Representation of Neural Networks Architectures

DGX agent

arXiv:2512.17593v3 Announce Type: replace Abstract: In this paper we consider the limiting case of neural networks (NNs) architectures when the number of neurons in each hidden layer and the number of

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Accelerating Power Method with Fast Sketching for Stronger Low-Rank Approximation

DGX agent

arXiv:2605.09755v1 Announce Type: cross Abstract: The power method is one of the most fundamental tools for extracting top principal components from data through low-rank matrix approximation. Yet, wh

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

AdamFLIP: Adaptive Momentum Feedback Linearization Optimization for Hard Constrained PINN Training

DGX agent

arXiv:2605.08408v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) provide a flexible framework for solving forward and inverse problems governed by partial differential equation

model-releasesarxiv-cs-lg
12 May 2026
Safety

Adversarial Attacks Against MLLMs via Progressive Resolution Processing and Adaptive Feature Alignment

DGX agent

arXiv:2605.09902v1 Announce Type: new Abstract: Adversarial perturbations can mislead Multimodal Large Language Models (MLLMs) recognize a benign image as a specific target object, posing serious risk

safetyarxiv-cs-cv
12 May 2026
Model Releases

Adversary-Robust Learning from Fully Asynchronous Directional Derivative Estimates

DGX agent

arXiv:2605.09337v1 Announce Type: new Abstract: We propose FAR-SIGN (Fully Asynchronous Robust optimization via SIGNed directional projections) for adversary-resilient learning in parameter-server--wo

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Agentic MIP Research: Accelerated Constraint Handler Generation

DGX agent

arXiv:2605.09186v1 Announce Type: new Abstract: Mixed-integer programming (MIP) research is both mathematically sophisticated and engineering-intensive: testing an algorithmic hypothesis within a bran

model-releasesarxiv-cs-ai
12 May 2026
Applications

AI Gateway production index

DGX agent

Vercel's AI Gateway production index is a monitoring tool that tracks the performance and reliability of AI services and models in production environments. It likely provides metrics on latency, uptim

applicationsvercel-blog
12 May 2026
Model Releases

AlphaExploitem: Going Beyond the Nash Equilibrium in Poker by Learning to Exploit Suboptimal Play

DGX agent

arXiv:2605.09150v1 Announce Type: new Abstract: Poker is an imperfect information game that has served as a long-standing benchmark for decision-making under uncertainty. To maximize utility beyond th

model-releasesarxiv-cs-lg
12 May 2026
Research

Amortizing Causal Sensitivity Analysis via Prior Data-Fitted Networks

DGX agent

arXiv:2605.10590v1 Announce Type: cross Abstract: Causal sensitivity analysis aims to provide bounds for causal effect estimates in the presence of unobserved confounding. However, existing methods fo

researcharxiv-cs-lg
12 May 2026
Model Releases

Anthropic announces 12 Claude plugins for the legal sector, including a 'commercial counsel' tool for reviewing vendor agreements and a bar exam study tool (Rachel Metz/Bloomberg)

DGX agent

Rachel Metz / Bloomberg: Anthropic announces 12 Claude plugins for the legal sector, including a “commercial counsel” tool for reviewing vendor agreements and a bar exam study tool — Anthropic PBC is

model-releasestechmeme
12 May 2026
Model Releases

Arcane: An Assertion Reduction Framework through Semantic Clustering and MCTS-Guided Rule Exploring

DGX agent

arXiv:2605.10107v1 Announce Type: new Abstract: Assertion-based Verification (ABV) is essential for ensuring that hardware designs conform to their intended specifications. However, existing automated

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

AssemPlanner: A Multi-Agent Based Task Planning Framework for Flexible Assembly System

DGX agent

arXiv:2605.08831v1 Announce Type: new Abstract: In flexible assembly systems, existing task planning methods require a time-consuming configuration process by multiple experts to establish a productio

model-releasesarxiv-cs-ro
12 May 2026
Model Releases

ASTRA-QA: A Benchmark for Abstract Question Answering over Documents

DGX agent

arXiv:2605.10168v1 Announce Type: new Abstract: Document-based question answering (QA) increasingly includes abstract questions that require synthesizing scattered information from long documents or a

model-releasesarxiv-cs-cl
12 May 2026
Research

Attractor-Vascular Coupling Theory: Formal Grounding and Empirical Validation for AAMI-Standard Cuffless Blood Pressure Estimation from Smartphone Photoplethysmography

DGX agent

arXiv:2605.10871v1 Announce Type: cross Abstract: This work proposes Attractor-Vascular Coupling Theory (AVCT), a mathematical framework showing that cardiac attractor geometry encodes blood pressure

researcharxiv-cs-ai
12 May 2026
← Previous
1…866867868869870…1312
Next →