AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,585 results
Model Releases

Vision-Based Hand Shadowing for Robotic Manipulation via Inverse Kinematics

DGX agent

arXiv:2603.11383v2 Announce Type: replace Abstract: Teleoperation of low-cost robotic manipulators remains challenging due to the difficulty of retargeting human hand motion to robot joint commands. W

model-releasesarxiv-cs-ro
13 May 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Vision2Code: A Multi-Domain Benchmark for Evaluating Image-to-Code Generation

DGX agent

arXiv:2605.11307v1 Announce Type: new Abstract: Image-to-code generation tests whether a vision-language model (VLM) can recover the structure of an image enough to express it as executable code. Exis

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

We built SmithDB: the database purpose built for agent observability workloads that now powers many parts of LangSmith. Agent observability …

DGX agent

We built SmithDB: the database purpose built for agent observability workloads that now powers many parts of LangSmith. Agent observability presents a challenging data problem. Agent traces can contai

model-releasesharrison-chase--x
13 May 2026
Model Releases

🎉 We published a new AI safety study: shopping agents fall for whimsical attacks and lose money. A whimsical attack is an absurd scenario a…

DGX agent

🎉 We published a new AI safety study: shopping agents fall for whimsical attacks and lose money. A whimsical attack is an absurd scenario a human would never try on another human. In one run, GPT-5.1

model-releasesemad-mostaque--x
13 May 2026
Model Releases

Weather-Robust Cross-View Geo-Localization via Prototype-Based Semantic Part Discovery

DGX agent

arXiv:2605.11654v1 Announce Type: new Abstract: Cross-view geo-localization (CVGL), which matches an oblique drone view to a geo-referenced satellite tile, has emerged as a key alternative for autonom

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

What Does It Mean for a Medical AI System to Be Right?

DGX agent

arXiv:2605.11963v1 Announce Type: new Abstract: This paper examines what it means for a medical AI system to be right by grounding the question in a specific clinical context: the automatic classifica

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

When you want to move from single agent SQLite on something like QMD, PostgreSQL is a great choice for multi agent and production quality, b…

DGX agent

When you want to move from single agent SQLite on something like QMD, PostgreSQL is a great choice for multi agent and production quality, but not as snappy. So we made it much more snappy with BM25 &

model-releasesemad-mostaque--x
13 May 2026
Model Releases

💡 Why did @togethercompute choose NVIDIA Blackwell to serve DeepSeek-V4? Because NVIDIA Blackwell is built for the bottlenecks that matter …

DGX agent

💡 Why did @togethercompute choose NVIDIA Blackwell to serve DeepSeek-V4? Because NVIDIA Blackwell is built for the bottlenecks that matter most in long-context inference: → KV-cache pressure during de

model-releasestogether-ai--x
13 May 2026
Model Releases

WildRelight: A Real-World Benchmark and Physics-Guided Adaptation for Single-Image Relighting

DGX agent

arXiv:2605.11696v1 Announce Type: new Abstract: Recent single-image relighting methods, powered by advanced generative models, have achieved impressive photorealism on synthetic benchmarks. However, t

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Will Ollama come out with a non-cloud version of Deepseek-v4 Flash?

DGX agent

DeepSeek-v4 Flash through Ollama is currently available as a cloud model, where Ollama's CLI sends API calls to Ollama's hosted version rather than running locally . Local support for DeepSeek V4 Flas

model-releasesr-ollama
13 May 2026
Model Releases

XWOD: A Real-World Benchmark for Object Detection under Extreme Weather Conditions

DGX agent

arXiv:2605.11521v1 Announce Type: new Abstract: Autonomous driving and intelligent transportation systems remain vulnerable under extreme weather. The U.S. Federal Highway Administration reports that

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

YFPO: A Preliminary Study of Yoked Feature Preference Optimization with Neuron-Guided Rewards for Mathematical Reasoning

DGX agent

arXiv:2605.11906v1 Announce Type: new Abstract: Preference optimization has become an important post-training paradigm for improving the reasoning abilities of large language models. Existing methods

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

100,000+ Movie Reviews from Kazakhstan: Russian, Kazakh, and Code-Switched Texts

DGX agent

arXiv:2605.08600v1 Announce Type: new Abstract: We present a new publicly available corpus of 100,502 movie reviews from Kazakhstan collected from kino.kz, spanning 2001-2025 and covering 4,943 unique

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

3DReflecNet: A Large-Scale Dataset for 3D Reconstruction of Reflective, Transparent, and Low-Texture Objects

DGX agent

arXiv:2605.10204v1 Announce Type: new Abstract: Accurate 3D reconstruction of objects with reflective, transparent, or low-texture surfaces still remains notoriously challenging. Such materials often

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

A Cognitively Grounded Bayesian Framework for Misinformation Susceptibility

DGX agent

arXiv:2605.09483v1 Announce Type: cross Abstract: In this (work in progress) paper, we present Bounded Pragmatic Listener (or BPL), a cognitively grounded Bayesian framework for modelling susceptibili

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

A Communication-Theoretic Framework for LLM Agents: Cost-Aware Adaptive Reliability

DGX agent

arXiv:2605.09121v1 Announce Type: cross Abstract: Agents built on large language models (LLMs) rely on a range of reliability techniques, including retry, majority voting, and self-consistency, that h

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

A Deep Risk Estimator for Known Operator Learning

DGX agent

arXiv:2605.08517v1 Announce Type: cross Abstract: We describe an approach for estimating the statistical risk of deep networks that contain a mix of learned and known operators. Building on the maxima

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

A Game Theoretic Free Energy Analysis of Higher Order Synergy in Attention Heads of Large Language Models

DGX agent

arXiv:2605.09515v1 Announce Type: new Abstract: Large language models rely on multihead attention, but interactions among heads remain poorly understood. We apply the Game Theoretic Free Energy Princi

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

A Geometric Perspective on Next-Token Prediction in Large Language Models: Three Emerging Phases

DGX agent

arXiv:2605.09011v1 Announce Type: cross Abstract: We investigate the geometry of predictive information across the layers of large language models (LLMs). We repurpose representation lenses-learned af

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

A meshfree exterior calculus for generalizable and data-efficient learning of physics from point clouds

DGX agent

arXiv:2605.08436v1 Announce Type: cross Abstract: We introduce a meshfree exterior calculus (MEEC) for learning structure-preserving descriptions of physics on point clouds, and use it to build MEEC-N

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

A new initialisation to Control Gradients in Sinusoidal Neural network

DGX agent

arXiv:2512.06427v2 Announce Type: replace Abstract: Proper initialisation strategy is of primary importance to mitigate gradient explosion or vanishing when training neural networks. Yet, the impact o

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

A Quantum Inspired Variational Kernel and Explainable AI Framework for Cross Region Solar and Wind Energy Forecasting

DGX agent

arXiv:2605.09032v1 Announce Type: cross Abstract: Reliable short horizon forecasting of solar and wind generation is a structural prerequisite of any modern power system yet most published forecasters

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

A Stability Benchmark of Generative Regularizers for Inverse Problems

DGX agent

arXiv:2605.10076v1 Announce Type: cross Abstract: Generative (diffusion) priors demonstrate remarkable performance in addressing inverse problems in imaging. Yet, for scientific and medical imaging, i

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

A startup wants to pull magnesium from seawater without torching the environment. Another wants to take small language models to the podium …

DGX agent

A startup wants to pull magnesium from seawater without torching the environment. Another wants to take small language models to the podium for enterprise customers. Today @jason and @alex sat down wi

model-releasesai21-labs--x
12 May 2026
Model Releases

A Unified Representation of Neural Networks Architectures

DGX agent

arXiv:2512.17593v3 Announce Type: replace Abstract: In this paper we consider the limiting case of neural networks (NNs) architectures when the number of neurons in each hidden layer and the number of

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Accelerating Power Method with Fast Sketching for Stronger Low-Rank Approximation

DGX agent

arXiv:2605.09755v1 Announce Type: cross Abstract: The power method is one of the most fundamental tools for extracting top principal components from data through low-rank matrix approximation. Yet, wh

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Acceptance Cards:A Four-Diagnostic Standard for Safe Fine-Tuning Defense Claims

DGX agent

arXiv:2605.10575v1 Announce Type: cross Abstract: Safe fine-tuning defenses are often endorsed on the basis of a held-out gap reduction, but the same reduction can come from sampling noise, subject ar

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Action-Guided Attention for Video Action Anticipation

DGX agent

arXiv:2603.01743v2 Announce Type: replace Abstract: Anticipating future actions in videos is challenging, as the observed frames provide only evidence of past activities, requiring the inference of la

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

ACWM-Phys: Investigating Generalized Physical Interaction in Action-Conditioned Video World Models

DGX agent

arXiv:2605.08567v1 Announce Type: new Abstract: Action-conditioned world models (ACWMs) have shown strong promise for video prediction and decision-making. However, existing benchmarks are largely res

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

AdamFLIP: Adaptive Momentum Feedback Linearization Optimization for Hard Constrained PINN Training

DGX agent

arXiv:2605.08408v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) provide a flexible framework for solving forward and inverse problems governed by partial differential equation

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

AdaPaD: Adaptive Parallel Deflation for PEFT with Self-Correcting Rank Discovery

DGX agent

arXiv:2605.10741v1 Announce Type: new Abstract: Fine-tuning large language models with LoRA requires choosing a rank r before training starts. Existing approaches either extract rank-1 components sequ

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

AdaPreLoRA: Adafactor Preconditioned Low-Rank Adaptation

DGX agent

arXiv:2605.08734v1 Announce Type: cross Abstract: Low-Rank Adaptation (LoRA) reparameterizes a weight update as a product of two low-rank factors, but the Jacobian J_{G} of the generator mapping the f

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Adaptive Multi-view Graph Contrastive Learning via Fractional-order Neural Diffusion Networks

DGX agent

arXiv:2511.06216v4 Announce Type: replace Abstract: Graph contrastive learning (GCL) learns node and graph representations by contrasting multiple views of the same graph. Existing methods typically r

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Adversary-Robust Learning from Fully Asynchronous Directional Derivative Estimates

DGX agent

arXiv:2605.09337v1 Announce Type: new Abstract: We propose FAR-SIGN (Fully Asynchronous Robust optimization via SIGNed directional projections) for adversary-resilient learning in parameter-server--wo

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Agent-ValueBench: A Comprehensive Benchmark for Evaluating Agent Values

DGX agent

arXiv:2605.10365v1 Announce Type: new Abstract: Autonomous agents have rapidly matured as task executors and seen widespread deployment via harnesses such as OpenClaw. Safety concerns have rightly dra

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

AgentCollabBench: Diagnosing When Good Agents Make Bad Collaborators

DGX agent

arXiv:2605.08647v1 Announce Type: cross Abstract: Multi-agent systems achieve state-of-the-art outcomes through peer collaboration. However, when an agent in the pipeline silently drops a constraint,

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

AgentForesight: Online Auditing for Early Failure Prediction in Multi-Agent Systems

DGX agent

arXiv:2605.08715v1 Announce Type: cross Abstract: LLM-based multi-agent systems are increasingly deployed on long-horizon tasks, but a single decisive error is often accepted by downstream agents and

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Agentic MIP Research: Accelerated Constraint Handler Generation

DGX agent

arXiv:2605.09186v1 Announce Type: new Abstract: Mixed-integer programming (MIP) research is both mathematically sophisticated and engineering-intensive: testing an algorithmic hypothesis within a bran

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Agentic Performance at the Edge: Insights from Benchmarking

DGX agent

arXiv:2605.10384v1 Announce Type: new Abstract: Agentic artificial intelligence (AI) is a natural fit for Internet of Things (IoT) and edge systems, but edge deployments are often constrained to model

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

AgentPSO: Evolving Agent Reasoning Skill via Multi-agent Particle Swarm Optimization

DGX agent

arXiv:2605.08704v1 Announce Type: new Abstract: Multi-agent reasoning has shown promise for improving the problem-solving ability of large language models by allowing multiple agents to explore divers

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks

DGX agent

arXiv:2605.10286v1 Announce Type: new Abstract: Building effective clinical decision support systems requires the synthesis of complex heterogeneous multimodal data. Such modalities include temporal e

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

AHD Agent: Agentic Reinforcement Learning for Automatic Heuristic Design

DGX agent

arXiv:2605.08756v1 Announce Type: new Abstract: Automatic heuristic design (AHD) has emerged as a promising paradigm for solving NP-hard combinatorial optimization problems (COPs). Recent works show t

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

AI has a secondary intent problem. I was negotiating a rent renewal (I live in NYC and they raised rent 10% because I guess they were bored)…

DGX agent

AI has a secondary intent problem. I was negotiating a rent renewal (I live in NYC and they raised rent 10% because I guess they were bored). The property manager said she would 'do everything she cou

model-releasesallie-k--miller--x
12 May 2026
Model Releases

Alice v1: Distillation-Enhanced Video Generation Surpassing Closed-Source Models

DGX agent

arXiv:2605.08115v1 Announce Type: cross Abstract: Wepresent Alice v1, a 14-billion parameter open-source video generation model that achieves state-of-the-art quality through consistency distillation

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

AlignDrive: Aligned Lateral-Longitudinal Planning for End-to-End Autonomous Driving

DGX agent

arXiv:2601.01762v2 Announce Type: replace-cross Abstract: Practical autonomous driving requires models that generalize by reasoning through spatial-temporal possibilities to exclude unsafe outcomes. W

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Aligning Agents via Planning: A Benchmark for Trajectory-Level Reward Modeling

DGX agent

arXiv:2604.08178v2 Announce Type: replace Abstract: In classical Reinforcement Learning from Human Feedback (RLHF), Reward Models (RMs) serve as the fundamental signal provider for model alignment. As

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

AlphaExploitem: Going Beyond the Nash Equilibrium in Poker by Learning to Exploit Suboptimal Play

DGX agent

arXiv:2605.09150v1 Announce Type: new Abstract: Poker is an imperfect information game that has served as a long-standing benchmark for decision-making under uncertainty. To maximize utility beyond th

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

AlpsBench: An LLM Personalization Benchmark for Real-Dialogue Memorization and Preference Alignment

DGX agent

arXiv:2603.26680v2 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) evolve into lifelong AI assistants, LLM personalization has become a critical frontier. However, progress is c

model-releasesarxiv-cs-ai
12 May 2026
← Previous
1…324325326327328…471
Next →