AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
87,814Total entries
1Added by human
87,813Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,137 results
Local Ai

TextSeal: A Localized LLM Watermark for Provenance & Distillation Protection

DGX agent

arXiv:2605.12456v1 Announce Type: cross Abstract: We introduce TextSeal, a state-of-the-art watermark for large language models. Building on Gumbel-max sampling, TextSeal introduces dual-key generatio

local-aiarxiv-cs-cl
13 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

The Challenge and Reward of Fair Play in Narrative: A Computational Approach

DGX agent

arXiv:2507.13841v2 Announce Type: replace Abstract: Good storytelling involves surprise -- unpredictability in how the story unfolds -- and sense-making, the requirement that the story forms a coheren

researcharxiv-cs-cl
13 May 2026
Model Releases

The UK’s state AI Security iIstitute findings: 1) Mythos is a big gain in cyber capabilities. But so is GPT-5.5 2) It is hard to establish a…

DGX agent

The UK’s state AI Security iIstitute findings: 1) Mythos is a big gain in cyber capabilities. But so is GPT-5.5 2) It is hard to establish an upper bound on Mythos/GPT-5.5, which appear to be limited

model-releasesethan-mollick--x
13 May 2026
Model Releases

This is misleading. This policy redefines the term 'interactive' to mean 'using an Anthropic front-end'. If you use `claude -p` or Agent SDK…

DGX agent

This is misleading. This policy redefines the term 'interactive' to mean 'using an Anthropic front-end'. If you use `claude -p` or Agent SDK to do something interactively, it now uses credits, not you

model-releasesjeremy-howard--x
13 May 2026
Model Releases

This jackass fought hard to remain NASA Administrator.

DGX agent

This jackass fought hard to remain NASA Administrator. SCOOP: I obtained a pitch deck in which the entity that paid for Transportation Sec’y Duffy’s new reality show outlined different partner levels

model-releasesanthropic--x
13 May 2026
Model Releases

TikTok launches TikTok GO in the US for users to book hotels, attractions, and experiences directly in the app, partnering with Booking.com, Expedia, and others (Aisha Malik/TechCrunch)

DGX agent

Aisha Malik / TechCrunch: TikTok launches TikTok GO in the US for users to book hotels, attractions, and experiences directly in the app, partnering with Booking.com, Expedia, and others — TikTok anno

model-releasestechmeme
13 May 2026
Safety

TMPO: Trajectory Matching Policy Optimization for Diverse and Efficient Diffusion Alignment

DGX agent

arXiv:2605.10983v1 Announce Type: cross Abstract: Reinforcement learning (RL) has shown extraordinary potential in aligning diffusion models to downstream tasks, yet most of them still suffer from sig

safetyarxiv-cs-cv
13 May 2026
Safety

Towards Fine-Grained Code-Switch Speech Translation with Semantic Space Alignment

DGX agent

arXiv:2511.10670v2 Announce Type: replace Abstract: Code-switching (CS) speech translation (ST) aims to translate speech that alternates between multiple languages into a target language text, posing

safetyarxiv-cs-cl
13 May 2026
Research

Training-Inference Consistent Segmented Execution for Long-Context LLMs

DGX agent

arXiv:2605.11744v1 Announce Type: new Abstract: Transformer-based large language models face severe scalability challenges in long-context generation due to the computational and memory costs of full-

researcharxiv-cs-cl
13 May 2026
Model Releases

Trajectory-Agnostic Asteroid Detection in TESS with Deep Learning

DGX agent

arXiv:2605.12391v1 Announce Type: cross Abstract: We present a novel method for extracting moving objects from TESS data using machine learning. Our approach uses two stacked 3D U-Nets with skip conne

model-releasesarxiv-cs-lg
13 May 2026
Safety

Understanding and Preventing Entropy Collapse in RLVR with On-Policy Entropy Flow Optimization

DGX agent

arXiv:2605.11491v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has become an effective paradigm for improving the reasoning ability of large language models. How

safetyarxiv-cs-lg
13 May 2026
Safety

Understanding the Performance Gap in Preference Learning: A Dichotomy of RLHF and DPO

DGX agent

arXiv:2505.19770v5 Announce Type: replace-cross Abstract: We present a fine-grained theoretical analysis of the performance gap between two-stage reinforcement learning from human feedback~(RLHF) and

safetyarxiv-cs-cl
13 May 2026
Research

Uniform Scaling Limits in AdamW-Trained Transformers

DGX agent

arXiv:2605.11059v1 Announce Type: cross Abstract: We study the large-depth limit of transformers trained with AdamW, by modelling the hidden-state dynamics as an interacting particle system (IPS) coup

researcharxiv-cs-lg
13 May 2026
Applications

UniVLR: Unifying Text and Vision in Visual Latent Reasoning for Multimodal LLMs

DGX agent

arXiv:2605.11856v1 Announce Type: cross Abstract: Multimodal large language models are increasingly expected to perform thinking with images, yet existing visual latent reasoning methods still rely on

applicationsarxiv-cs-cl
13 May 2026
Local Ai

v0.23.4

DGX agent

Ollama v0.23.4 is a release version of Ollama, a tool for running large language models locally. This patch release likely includes bug fixes, performance improvements, and refinements to existing fea

local-aiollama-releases
13 May 2026
Model Releases

Very Efficient Listwise Multimodal Reranking for Long Documents

DGX agent

arXiv:2605.11864v1 Announce Type: cross Abstract: Listwise reranking is a key yet computationally expensive component in vision-centric retrieval and multimodal retrieval-augmented generation (M-RAG)

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Vision-Based Hand Shadowing for Robotic Manipulation via Inverse Kinematics

DGX agent

arXiv:2603.11383v2 Announce Type: replace Abstract: Teleoperation of low-cost robotic manipulators remains challenging due to the difficulty of retargeting human hand motion to robot joint commands. W

model-releasesarxiv-cs-ro
13 May 2026
Model Releases

We built SmithDB: the database purpose built for agent observability workloads that now powers many parts of LangSmith. Agent observability …

DGX agent

We built SmithDB: the database purpose built for agent observability workloads that now powers many parts of LangSmith. Agent observability presents a challenging data problem. Agent traces can contai

model-releasesharrison-chase--x
13 May 2026
Model Releases

Weather-Robust Cross-View Geo-Localization via Prototype-Based Semantic Part Discovery

DGX agent

arXiv:2605.11654v1 Announce Type: new Abstract: Cross-view geo-localization (CVGL), which matches an oblique drone view to a geo-referenced satellite tile, has emerged as a key alternative for autonom

model-releasesarxiv-cs-cv
13 May 2026
Research

When Looking Is Not Enough: Visual Attention Structure Reveals Hallucination in MLLMs

DGX agent

arXiv:2605.11559v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have become a key interface for visual reasoning and grounded question answering, yet they remain vulnerable to

researcharxiv-cs-cv
13 May 2026
Model Releases

When you want to move from single agent SQLite on something like QMD, PostgreSQL is a great choice for multi agent and production quality, b…

DGX agent

When you want to move from single agent SQLite on something like QMD, PostgreSQL is a great choice for multi agent and production quality, but not as snappy. So we made it much more snappy with BM25 &

model-releasesemad-mostaque--x
13 May 2026
Model Releases

💡 Why did @togethercompute choose NVIDIA Blackwell to serve DeepSeek-V4? Because NVIDIA Blackwell is built for the bottlenecks that matter …

DGX agent

💡 Why did @togethercompute choose NVIDIA Blackwell to serve DeepSeek-V4? Because NVIDIA Blackwell is built for the bottlenecks that matter most in long-context inference: → KV-cache pressure during de

model-releasestogether-ai--x
13 May 2026
Model Releases

XWOD: A Real-World Benchmark for Object Detection under Extreme Weather Conditions

DGX agent

arXiv:2605.11521v1 Announce Type: new Abstract: Autonomous driving and intelligent transportation systems remain vulnerable under extreme weather. The U.S. Federal Highway Administration reports that

model-releasesarxiv-cs-cv
13 May 2026
Research

10 days wasn't enough. Step 3.5 Flash⚡ is back on @NousResearch Portal free for the next 15 days!

DGX agent

Nous Research has made Step 3.5 Flash, a faster AI model variant, available again on their portal at no cost for a 15-day period following popular demand after an initial 10-day availability window. T

researchnous-research--x
12 May 2026
Model Releases

A Deep Risk Estimator for Known Operator Learning

DGX agent

arXiv:2605.08517v1 Announce Type: cross Abstract: We describe an approach for estimating the statistical risk of deep networks that contain a mix of learned and known operators. Building on the maxima

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

A meshfree exterior calculus for generalizable and data-efficient learning of physics from point clouds

DGX agent

arXiv:2605.08436v1 Announce Type: cross Abstract: We introduce a meshfree exterior calculus (MEEC) for learning structure-preserving descriptions of physics on point clouds, and use it to build MEEC-N

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

A new initialisation to Control Gradients in Sinusoidal Neural network

DGX agent

arXiv:2512.06427v2 Announce Type: replace Abstract: Proper initialisation strategy is of primary importance to mitigate gradient explosion or vanishing when training neural networks. Yet, the impact o

model-releasesarxiv-cs-lg
12 May 2026
Agents

A Prompt-Aware Structuring Framework for Reliable Reuse of AI-Generated Content in the Agentic Web

DGX agent

arXiv:2605.09283v1 Announce Type: new Abstract: The evolution of Large Language Models (LLMs) and the software agents built on them (AI agents) marks a turning point in the transition from a human-cen

agentsarxiv-cs-ai
12 May 2026
Research

A PyTorch Library of Turing-Complete Neural Networks

DGX agent

arXiv:2605.08150v1 Announce Type: new Abstract: We present a PyTorch package that compiles neural networks and their weights from Turing machine descriptions, producing models that exactly simulate th

researcharxiv-cs-lg
12 May 2026
Model Releases

A Unified Representation of Neural Networks Architectures

DGX agent

arXiv:2512.17593v3 Announce Type: replace Abstract: In this paper we consider the limiting case of neural networks (NNs) architectures when the number of neurons in each hidden layer and the number of

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Accelerating Power Method with Fast Sketching for Stronger Low-Rank Approximation

DGX agent

arXiv:2605.09755v1 Announce Type: cross Abstract: The power method is one of the most fundamental tools for extracting top principal components from data through low-rank matrix approximation. Yet, wh

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

AdamFLIP: Adaptive Momentum Feedback Linearization Optimization for Hard Constrained PINN Training

DGX agent

arXiv:2605.08408v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) provide a flexible framework for solving forward and inverse problems governed by partial differential equation

model-releasesarxiv-cs-lg
12 May 2026
Safety

Adversarial Attacks Against MLLMs via Progressive Resolution Processing and Adaptive Feature Alignment

DGX agent

arXiv:2605.09902v1 Announce Type: new Abstract: Adversarial perturbations can mislead Multimodal Large Language Models (MLLMs) recognize a benign image as a specific target object, posing serious risk

safetyarxiv-cs-cv
12 May 2026
Model Releases

Adversary-Robust Learning from Fully Asynchronous Directional Derivative Estimates

DGX agent

arXiv:2605.09337v1 Announce Type: new Abstract: We propose FAR-SIGN (Fully Asynchronous Robust optimization via SIGNed directional projections) for adversary-resilient learning in parameter-server--wo

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Agentic MIP Research: Accelerated Constraint Handler Generation

DGX agent

arXiv:2605.09186v1 Announce Type: new Abstract: Mixed-integer programming (MIP) research is both mathematically sophisticated and engineering-intensive: testing an algorithmic hypothesis within a bran

model-releasesarxiv-cs-ai
12 May 2026
Applications

AI Gateway production index

DGX agent

Vercel's AI Gateway production index is a monitoring tool that tracks the performance and reliability of AI services and models in production environments. It likely provides metrics on latency, uptim

applicationsvercel-blog
12 May 2026
Model Releases

AlphaExploitem: Going Beyond the Nash Equilibrium in Poker by Learning to Exploit Suboptimal Play

DGX agent

arXiv:2605.09150v1 Announce Type: new Abstract: Poker is an imperfect information game that has served as a long-standing benchmark for decision-making under uncertainty. To maximize utility beyond th

model-releasesarxiv-cs-lg
12 May 2026
Research

Amortizing Causal Sensitivity Analysis via Prior Data-Fitted Networks

DGX agent

arXiv:2605.10590v1 Announce Type: cross Abstract: Causal sensitivity analysis aims to provide bounds for causal effect estimates in the presence of unobserved confounding. However, existing methods fo

researcharxiv-cs-lg
12 May 2026
Model Releases

Anthropic announces 12 Claude plugins for the legal sector, including a 'commercial counsel' tool for reviewing vendor agreements and a bar exam study tool (Rachel Metz/Bloomberg)

DGX agent

Rachel Metz / Bloomberg: Anthropic announces 12 Claude plugins for the legal sector, including a “commercial counsel” tool for reviewing vendor agreements and a bar exam study tool — Anthropic PBC is

model-releasestechmeme
12 May 2026
Model Releases

Arcane: An Assertion Reduction Framework through Semantic Clustering and MCTS-Guided Rule Exploring

DGX agent

arXiv:2605.10107v1 Announce Type: new Abstract: Assertion-based Verification (ABV) is essential for ensuring that hardware designs conform to their intended specifications. However, existing automated

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

AssemPlanner: A Multi-Agent Based Task Planning Framework for Flexible Assembly System

DGX agent

arXiv:2605.08831v1 Announce Type: new Abstract: In flexible assembly systems, existing task planning methods require a time-consuming configuration process by multiple experts to establish a productio

model-releasesarxiv-cs-ro
12 May 2026
Model Releases

ASTRA-QA: A Benchmark for Abstract Question Answering over Documents

DGX agent

arXiv:2605.10168v1 Announce Type: new Abstract: Document-based question answering (QA) increasingly includes abstract questions that require synthesizing scattered information from long documents or a

model-releasesarxiv-cs-cl
12 May 2026
Research

Attractor-Vascular Coupling Theory: Formal Grounding and Empirical Validation for AAMI-Standard Cuffless Blood Pressure Estimation from Smartphone Photoplethysmography

DGX agent

arXiv:2605.10871v1 Announce Type: cross Abstract: This work proposes Attractor-Vascular Coupling Theory (AVCT), a mathematical framework showing that cardiac attractor geometry encodes blood pressure

researcharxiv-cs-ai
12 May 2026
Model Releases

Automated Approach for Solving Infinite-state Polynomial Reachability Games

DGX agent

arXiv:2605.10169v1 Announce Type: new Abstract: Reachability games are two-player games played on a graph, where the objective of exttt{REACH} player is to reach the target set whereas the objective o

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

BabelDOC: Better Layout-Preserving PDF Translation via Intermediate Representation

DGX agent

arXiv:2605.10845v1 Announce Type: cross Abstract: As global cross-lingual communication intensifies, language barriers in visually rich documents such as PDFs remain a practical bottleneck. Existing d

model-releasesarxiv-cs-cl
12 May 2026
Agents

Batch-of-Thought: Cross-Instance Learning for Enhanced LLM Reasoning

DGX agent

arXiv:2601.02950v3 Announce Type: replace Abstract: Current Large Language Model reasoning systems process queries independently, discarding valuable cross-instance signals such as shared reasoning pa

agentsarxiv-cs-ai
12 May 2026
Model Releases

BCJR-QAT: A Differentiable Relaxation of Trellis-Coded Weight Quantization

DGX agent

arXiv:2605.10655v1 Announce Type: new Abstract: Trellis-coded quantization sets the current 2-bit post-training frontier for LLMs (QTIP), but pushing below the PTQ ceiling requires quantization-aware

model-releasesarxiv-cs-lg
12 May 2026
Local Ai

Beyond Bag-of-Patches: Learning Global Layout via Textual Supervision for Late-Interaction Visual Document Retrieval

DGX agent

arXiv:2605.08421v1 Announce Type: new Abstract: Visual Document Retrieval (VDR) models mostly rely on late interaction architectures, in which documents are represented by a set of local patch embeddi

local-aiarxiv-cs-cv
12 May 2026
← Previous
1…869870871872873…1316
Next →