AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
All
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,323 results
Model Releases

Large Causal Models for Temporal Causal Discovery

DGX agent

arXiv:2602.18662v2 Announce Type: replace Abstract: Causal discovery for both cross-sectional and temporal data has traditionally followed a dataset-specific paradigm, where a new model is fitted for

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Latency-Tolerant Cloud-Edge Collaborative Vision-Language-Action Models via Emergent Representational Specialization

DGX agent

arXiv:2608.00569v1 Announce Type: new Abstract: Deploying billion-parameter Vision-Language-Action (VLA) policies on mobile robots creates a systems conflict: semantic reasoning benefits from cloud GP

model-releasesarxiv-cs-ro
4 Aug 2026
Model Releases

Latent-Centroid Steering: Single-Pass Classifier-Free Guidance for Command-Aligned Autonomous Driving

DGX agent

arXiv:2608.00237v1 Announce Type: new Abstract: Vision-language models (VLMs) have recently emerged as a promising paradigm for end-to-end autonomous driving, enabling agents to map multimodal inputs

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Learning Compositional Meta-Routing for Agentic Workflows: An Executable Benchmark

DGX agent

arXiv:2608.00106v1 Announce Type: new Abstract: Agentic systems must decide not only what answer to produce, but which reasoning and execution operations should precede it. A controller may answer dir

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Learning What to Remember: Test-Time Training via Context Distillation

DGX agent

arXiv:2608.01672v1 Announce Type: new Abstract: Effective long-context modeling is not merely about retaining more of the past, but about preserving the information that may prove relevant later. Test

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Lethe: How Hard Is It to Forget? A Benchmark for Federated Unlearning in Medical Imaging

DGX agent

arXiv:2608.01094v1 Announce Type: new Abstract: Federated learning enables medical-imaging models to be trained across hospitals, and privacy law, most explicitly the GDPR ``right to be forgotten'', t

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

LexisNexis opens customer innovation lab driven by AI to change the future of legal work

DGX agent

LexisNexis Legal & Professional, a division of RELX plc, today announced the opening of its Customer Innovation Lab in New York City, which will deliver a new model for how legal artificial intelligen

model-releasessiliconangle
4 Aug 2026
Model Releases

LFM2.5-2.6B is out

DGX agent

Released today, with emphasis on agentic capabilities. I really like their models for simple, high volume tasks ('summarize these gazillion documents') and their 8b-a1b was my go-to for certain tasks

model-releasesr-localllama
4 Aug 2026
Model Releases

Light Forcing: Accelerating Autoregressive Video Diffusion via Sparse Attention

DGX agent

arXiv:2602.04789v4 Announce Type: replace Abstract: Advanced autoregressive (AR) video generation models have improved visual fidelity and interactivity, but the quadratic complexity of attention rema

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Linear Multi-Timescale Retention as a Memory-Efficient Vision-Language Bridge

DGX agent

arXiv:2608.01614v1 Announce Type: new Abstract: Vision-Language Models (VLMs) face a critical computational bottleneck when processing high-resolution imagery due to the O(N^2) memory complexity of So

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Live and ready to build. Thanks for having us! @OpenRouter. Open weights dropping soon.⚡️

DGX agent

Live and ready to build. Thanks for having us! @OpenRouter. Open weights dropping soon.⚡️ Qwen3.8 Max by @Alibaba_Qwen is live on OpenRouter. The new flagship has 2.4T parameters (95B active) and is b

model-releasesqwen--x
4 Aug 2026
Model Releases

LiveMem: Maintaining Memory State Continuity in Long-Running LLM Inference

DGX agent

arXiv:2608.02515v1 Announce Type: new Abstract: Long-running assistants and agents consume interaction streams that eventually outgrow the context. Existing context retention, summarization, and retri

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Llama.cpp PR 8% speed boost

DGX agent

Llama.cpp currently uses cpu based sampling for user with mtp enabled. The PR moves sampling to the gpu, which on a 5090 boasts an 8% increase in tok/s for qwen3.6:35b. I tested it on my P40 and obser

model-releasesr-localllama
4 Aug 2026
Model Releases

llm-anthropic 0.26

DGX agent

Release: llm-anthropic 0.26 Includes new features enabled by LLM 0.32: New models: claude-fable-5, claude-sonnet-5, and claude-opus-5. #75, #76 Added server-side tools for WebSearch, WebFetch, CodeExe

model-releasessimon-willison
4 Aug 2026
Model Releases

Loggia dei Lanzi: AI Thermography Enhancement Comparisons through 3D Photogrammetry

DGX agent

arXiv:2608.02404v1 Announce Type: new Abstract: The Loggia dei Lanzi in the Piazza della Signoria is one of Florence's most prominent structures visited by millions every year. Its construction histor

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Long-Horizon Embodied Decision-Making via Multimodal Memory Compression

DGX agent

arXiv:2608.01456v1 Announce Type: cross Abstract: Agents are increasingly expected to act not only as task executors, but also as decision-makers on behalf of human users. This shift requires agents t

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

LongCat Sparse Attention: Taming the Lightning via Streaming-aware Hierarchical Cross-Layer Indexing

DGX agent

arXiv:2608.01662v1 Announce Type: cross Abstract: DeepSeek Sparse Attention (DSA) enables efficient long-context modeling through its Lightning Indexer. However, practical deployment remains constrain

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

LongChart VQA: A Comprehensive Benchmark for MLLMs with Complex Multi-Chart Reasoning

DGX agent

arXiv:2608.01328v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) are rapidly evolving with expanded context windows and stronger reasoning capabilities, enabling multi-chart un

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

LongHorizon-Harness: Advancing Long-Horizon Agents for Real-World Tasks

DGX agent

arXiv:2608.01964v1 Announce Type: new Abstract: Large language model (LLM) agents increasingly undertake long-horizon tasks that require sustained reasoning, tool use, and revision across many interde

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Look Up and Look Back: Hidden Attention and Latent Orientation in a Frozen Foundation Model for Panoramic SLAM

DGX agent

arXiv:2608.00925v1 Announce Type: new Abstract: Monocular panoramic SLAM benefits from substantial visual overlap under large camera rotations, yet remains prone to errors caused by camera tilt, scale

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models

DGX agent

arXiv:2608.02197v1 Announce Type: new Abstract: Visual representations of VLA models remain unreliable for spatially precise robotic manipulation. We uncover that vision encoders in VLAs also exhibit

model-releasesarxiv-cs-ro
4 Aug 2026
Model Releases

Loop-Mamba: A Loop Mamba with Degradation-Aware and Shared Memory for Old Photo Restoration

DGX agent

arXiv:2608.02346v1 Announce Type: new Abstract: Old photographs often suffer from multiple coupled degradations, including scratches, cracks, fading, blur, noise, and missing regions, severely degradi

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

LooperMuscle: Fast and Stable Learning of Humanoid Whole-Body Tracking via Structured Mixture-of-Experts

DGX agent

arXiv:2608.00820v1 Announce Type: new Abstract: FastSAC-style methods significantly reduce humanoid motion training time but often suffer from notable performance degradation compared with PPO in whol

model-releasesarxiv-cs-ro
4 Aug 2026
Model Releases

LoopsBench: From Harness Engineering to Loop Engineering in Benchmarking Coding Agent

DGX agent

arXiv:2608.00267v1 Announce Type: cross Abstract: Coding agent infrastructure is shifting from harness engineering toward loop engineering as coding agents are deployed for sustained long-horizon soft

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Machine-Precision Prediction of Low-Dimensional Chaotic Systems from Noise-Free Data

DGX agent

arXiv:2507.09652v2 Announce Type: replace-cross Abstract: Low-dimensional chaotic systems such as the Lorenz-63 model are commonly used to benchmark system-agnostic methods for learning dynamics from

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Mamba Policy: Towards Efficient 3D Diffusion Policy with Hybrid Selective State Models

DGX agent

arXiv:2409.07163v3 Announce Type: replace-cross Abstract: Diffusion models have been widely employed in the field of 3D manipulation due to their efficient capability to learn distributions, allowing

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Manifold-GS: Certified Hybrid Assets via Varifold-Conservative Gaussian Splatting

DGX agent

arXiv:2608.00214v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) gives high-quality novel-view synthesis, but its adaptive radiance primitives are not directly usable as structured assets:

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

MAPLE: Metadata Augmented Private Language Evolution

DGX agent

arXiv:2603.19258v2 Announce Type: replace Abstract: Differentially private (DP) fine-tuning of large language models (LLMs) requires massive compute and full model access, which rules out state-of-the

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

MDTD-ArtIR: Benchmarking Image Editing and Restoration Models for Art Image Restoration under Texture-Overlay Degradations

DGX agent

arXiv:2608.00736v1 Announce Type: new Abstract: Restoring severely degraded visual media still remains a formidable challenge, as existing methods often hallucinate unnatural textures and contents, st

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

MDWD: A Street-Level Dataset for Municipal Solid Waste Detection in Dense Urban Environments

DGX agent

arXiv:2608.00257v1 Announce Type: new Abstract: Automated visual monitoring of urban environments is a growing Computer Vision research area, but municipal solid waste detection remains under-represen

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Measuring in-context algorithmic reasoning in language models against an exact Bayes-optimal standard

DGX agent

arXiv:2608.01575v1 Announce Type: new Abstract: Whether large language models perform genuine algorithmic reasoning or mere pattern completion is hard to test, because most benchmarks lack a ground tr

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

MedPRESS: A Multi-turn Benchmark for Patient-Pressure-Induced Medical Sycophancy in LLMs

DGX agent

arXiv:2608.02520v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for health-related advice. Existing research measures their safety with static questions rather than

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

MetaRoute-Bench: Evaluating Meta-Decision Policies for Agentic Workflow Routing

DGX agent

arXiv:2608.00107v1 Announce Type: new Abstract: Agentic systems must repeatedly decide whether to answer directly, decompose a task, invoke a tool, execute code, delegate to a specialist, verify an in

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

MIEScore: Human-Aligned Evaluation for Multi-Source Image Editing

DGX agent

arXiv:2608.02059v1 Announce Type: new Abstract: Recent advances in unified multimodal models have significantly improved text-guided image editing abilities. In particular, models such as Nano-Banana-

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Minute-Scale Training for Microrobot Navigation

DGX agent

arXiv:2608.00854v1 Announce Type: new Abstract: Microrobots hold significant potential for various applications, where targeted navigation is a basic requirement. Deep reinforcement learning (DRL) has

model-releasesarxiv-cs-ro
4 Aug 2026
Model Releases

Mistral releases Shieldstral, a 3B multimodal safety classifier that it says matches models up to 7x its size on text safety, available under Apache 2.0 (Mistral AI Blog)

DGX agent

Mistral AI Blog: Mistral releases Shieldstral, a 3B multimodal safety classifier that it says matches models up to 7x its size on text safety, available under Apache 2.0 — Every product that ships a m

model-releasestechmeme
4 Aug 2026
Model Releases

Mitigating Backdoors via Decoy Shortcuts and Knowledge Decoupling

DGX agent

arXiv:2608.00732v1 Announce Type: cross Abstract: Backdoor attacks pose a serious threat to deep neural networks, especially when training relies on third-party data, allowing adversaries to inject ma

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

MixedComplementarityProblems.jl: A Fast, Batched, Open-Source Interior Point Solver for Mixed Complementarity Problems

DGX agent

arXiv:2608.00959v1 Announce Type: cross Abstract: Mixed complementarity problems (MCPs) arise as the first-order optimality conditions of nonlinear programs and noncooperative games, and provide a nat

model-releasesarxiv-cs-ro
4 Aug 2026
Model Releases

MoCRA: Mixture of Compositional Rank-1 Atoms for 4K All-in-One Video Restoration

DGX agent

arXiv:2608.01829v1 Announce Type: new Abstract: Real-world video arrives hazy, rainy, dark, or noisy, and a deployable restorer faces three demands at once: no degradation label, native 4K output, and

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Models as Tools: An Agentic Coordination Framework for Unified Multimodal Visual Tracking

DGX agent

arXiv:2608.00847v1 Announce Type: new Abstract: Most current visual trackers adopt a matching-based architecture trained exclusively on tracking datasets, whose performance gains depend heavily on the

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

MoRAL: Sensor-Grounded BEV Reasoning for Compact VLMs toward Edge-Oriented Autonomous Driving

DGX agent

arXiv:2608.02449v1 Announce Type: new Abstract: Deploying vision-language models (VLMs) for safety-critical spatial reasoning on resource-constrained autonomous driving platforms requires both compact

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Morphology Aware Reversible Semantic Tokenization and Hierarchical Word Composition for Tamil Language Models

DGX agent

arXiv:2608.01153v1 Announce Type: new Abstract: Statistical subword tokenizers can process arbitrary text, but their units need not align with lexical or grammatical structure. This is especially impo

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Move What Matters: Parameter-Efficient Domain Adaptation via Optimal Transport Flow for Collaborative Perception

DGX agent

arXiv:2602.11565v5 Announce Type: replace Abstract: Efficient domain adaptation remains a fundamental challenge for deploying multi-agent systems across diverse environments in Vehicle-to-Everything (

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Multiple result sets: How Database Migration Service automates SQL server to PostgreSQL translation

DGX agent

In the Medium blog post, 'From MARS to SETOF REFCURSOR: Migrating Multi-Result Stored Procedures to PostgreSQL,' we explored the fundamental architectural differences between SQL Server and PostgreSQL

model-releasesgoogle-cloud-ai
4 Aug 2026
Model Releases

New release of LLM adds support for reasoning traces, OpenAI Responses, server-side tools, and smarter logging

DGX agent

I released LLM 0.32 this morning, the most significant new version of LLM since the initial launch of the project. The new version includes support for visible reasoning traces, server-side provider t

model-releasessimon-willison
4 Aug 2026
Model Releases

New York Smells: A Large Multimodal Dataset for Olfaction

DGX agent

arXiv:2511.20544v2 Announce Type: replace Abstract: While olfaction is central to how animals perceive the world, this rich chemical sensory modality remains largely inaccessible to machines. One key

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Not the Dimension, the Norm: What Matters in Gradient-Free Weight Perturbation of Language Models

DGX agent

arXiv:2608.01624v1 Announce Type: new Abstract: Adapting a language model to a task no longer requires training all of its weights, and a line of parameter-efficient methods has driven the trainable c

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Nova: An End-to-End MLIR Compiler for Deep Learning

DGX agent

arXiv:2608.00029v1 Announce Type: cross Abstract: The performance of deep learning models at scale relies heavily on how effectively high-level mathematical operations are mapped to underlying physica

model-releasesarxiv-cs-lg
4 Aug 2026
← Previous
1…4647484950…466
Next →