AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
All
85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,778 results
Model Releases

PersistBench: When Should Long-Term Memories Be Forgotten by LLMs?

DGX agent

arXiv:2602.01146v2 Announce Type: replace Abstract: Conversational assistants are increasingly integrating long-term memory with large language models (LLMs). This persistence of memories, e.g., the u

model-releasesarxiv-cs-ai
4 Jun 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Physics-Informed Video Generation via Mixture-of-Experts Latent Alignment

DGX agent

arXiv:2606.04737v1 Announce Type: new Abstract: Large-scale video generation models have made remarkable progress in semantic consistency and visual quality, producing videos that are increasingly coh

model-releasesarxiv-cs-cv
4 Jun 2026
Model Releases

Plan, Watch, Recover: A Benchmark and Architectures for Proactive Procedural Assistance

DGX agent

arXiv:2606.04970v1 Announce Type: cross Abstract: We envision a proactive multi-modal assistant system which gives users real-time step-by-step guidance on a procedural task, autonomously deciding ext

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

PoliticsBench: Benchmarking Political Values in Large Language Models with Multi-Turn Roleplay

DGX agent

arXiv:2603.23841v2 Announce Type: replace-cross Abstract: While Large Language Models (LLMs) are increasingly used as primary sources of information, their potential for political bias may impact thei

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Prompt-Level Distillation: A Non-Parametric Alternative to Model Fine-Tuning for Efficient Reasoning

DGX agent

arXiv:2602.21103v2 Announce Type: replace Abstract: Advanced reasoning typically requires Chain-of-Thought prompting, which is accurate but incurs prohibitive latency and substantial test-time inferen

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

Proof-Carrying Agent Actions: Model-Agnostic Runtime Governance for Heterogeneous Agent Systems

DGX agent

arXiv:2606.04104v1 Announce Type: cross Abstract: Agent systems execute through runtimes with very different control points: local coding tools, framework SDKs, managed agent platforms, API gateways,

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Provably Reduced Sample Cost in Prior-Guided Hyperparameter Optimization

DGX agent

arXiv:2606.04866v1 Announce Type: new Abstract: Large-scale hyperparameter optimization (HPO) in automated machine learning (AutoML) consumes substantial computational resources, raising growing conce

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

Pseudospectral Bounds for Transient Amplification in Coupled Gradient Descent

DGX agent

arXiv:2606.04031v1 Announce Type: new Abstract: Coupled gradient descent--where the update of one parameter block depends on another--underlies bilevel optimization, two-time-scale stochastic approxim

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

Pulled the trigger today and switched 100% of Lindy traffic to DeepSeek v4, churning from Anthropic models. Saves us millions of $ and we're…

DGX agent

Pulled the trigger today and switched 100% of Lindy traffic to DeepSeek v4, churning from Anthropic models. Saves us millions of $ and we're actually seeing an *increase* in performance on many core u

model-releasesclem-delangue--x
4 Jun 2026
Model Releases

QO-Bench: Diagnosing Query-Operator-Preserving Retrieval over Typed Event Tuples

DGX agent

arXiv:2606.04646v1 Announce Type: cross Abstract: Many real-world questions over business, legal, and scientific corpora are natural-language versions of database-style queries over records latent in

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

QPredSGG: Hybrid Quantum Predicate Learning for Long-Tailed Scene Graph Generation

DGX agent

arXiv:2606.04689v1 Announce Type: cross Abstract: Scene Graph Generation (SGG) requires relational reasoning over objects and their interactions, but performance is often limited by severe long-tail p

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

Quantum entanglement provides a competitive advantage in adversarial games

DGX agent

arXiv:2603.10289v2 Announce Type: replace-cross Abstract: Whether uniquely quantum resources confer advantages in fully classical, competitive environments remains an open question. Competitive zero-s

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

QuBLAST: A Framework for Quantizing Large Language Models with Block-Level Compression Approach and Activation Scaling Strategy

DGX agent

arXiv:2606.04620v1 Announce Type: cross Abstract: LLMs have become the state-of-the-art algorithms for solving NLP tasks. However, they typically come at huge computational and memory costs, thus maki

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

RAMPART: Registry-based Agentic Memory with Priority-Aware Runtime Transformation

DGX agent

arXiv:2606.04628v1 Announce Type: new Abstract: RAMPART is a compile-time memory model and pure in-RAM block registry for LLM-based agents. Context assembly is a programmable runtime operation where c

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

RAVQ-HoloNet: Rate-Adaptive Vector-Quantized Hologram Compression

DGX agent

arXiv:2511.21035v2 Announce Type: replace Abstract: Holography offers significant potential for AR/VR applications. However, its adoption is limited by the high demand for data compression. Existing d

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

Reasoning over Boundaries: Enhancing Specification Alignment via Test-time Deliberation

DGX agent

arXiv:2509.14760v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly applied in diverse real-world scenarios, each governed by bespoke behavioral and safety specifications

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

Revisiting Vul-RAG: Reproducibility and Replicability of RAG-based Vulnerability Detection with Open-Weight Models

DGX agent

arXiv:2606.04739v1 Announce Type: cross Abstract: Large language models (LLMs) have shown strong potential for automated software vulnerability detection, particularly in retrieval-augmented generatio

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

RIDE: An Open Dataset and Benchmark for Train Delay Prediction

DGX agent

arXiv:2606.05070v1 Announce Type: new Abstract: Train delay prediction is an important problem for both passengers and railway operators, yet progress in the field remains difficult to assess due to t

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

Robotics startup Generalist, which released its GEN-1 model to complete short physical tasks in April, raised 400M led by Radical Ventures at a 2B valuation (Dina Bass/Bloomberg)

DGX agent

Dina Bass / Bloomberg: Robotics startup Generalist, which released its GEN-1 model to complete short physical tasks in April, raised 400M led by Radical Ventures at a 2B valuation — The company raised

model-releasestechmeme
4 Jun 2026
Model Releases

Robust Multi-view Clustering against Imperfect Information

DGX agent

arXiv:2606.04343v1 Announce Type: new Abstract: Real-world multi-view data always suffer from imperfect information problem, where the view-specific observations are absent (i.e., Incomplete Views, IV

model-releasesarxiv-cs-cv
4 Jun 2026
Model Releases

Rollout-Level Advantage-Prioritized Experience Replay for GRPO

DGX agent

arXiv:2606.04560v1 Announce Type: cross Abstract: Reinforcement learning from verifiable rewards with GRPO is a standard approach for post-training reasoning LLMs. It remains sample inefficient. Each

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Safety by narrow control has shown to fail many times. Need more transparency on the absolute frontier, and openness close behind.

DGX agent

Safety by narrow control has shown to fail many times. Need more transparency on the absolute frontier, and openness close behind. I found another API that offers claude-oceanus-v1-p the pricing and t

model-releasesclem-delangue--x
4 Jun 2026
Model Releases

Safety Under Scaffolding: How Evaluation Conditions Shape Measured Safety

DGX agent

arXiv:2603.10044v2 Announce Type: replace-cross Abstract: A safety score earned on a benchmark need not predict how the same model behaves once it is wrapped in an agentic scaffold the benchmark never

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

SAM 3D: 3Dfy Anything in Images

DGX agent

arXiv:2511.16624v2 Announce Type: replace-cross Abstract: We present SAM 3D, a generative model for visually grounded 3D object reconstruction, predicting geometry, texture, and layout from a single i

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Scaling AI Agents: A Step-by-Step Guide to Deploying ADK on GKE Autopilot

DGX agent

While building AI agents locally using Google’s Agent Development Kit (ADK) is an excellent way to prototype, production-ready agents require a robust, scalable infrastructure. For developers looking

model-releasesgoogle-cloud-ai
4 Jun 2026
Model Releases

Scene-Centric Unsupervised Video Panoptic Segmentation

DGX agent

arXiv:2606.04925v1 Announce Type: new Abstract: Video panoptic segmentation (VPS) aims to jointly detect, segment, and track all objects while partitioning the video into semantically consistent regio

model-releasesarxiv-cs-cv
4 Jun 2026
Model Releases

Self-Evolving Deep Research via Joint Generation and Evaluation

DGX agent

arXiv:2606.04507v1 Announce Type: cross Abstract: Large Language Models (LLMs) have become increasingly adopted in daily applications, with deep research standing out as a particularly important capab

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Shifting the Breaking Point of Flow Matching for Multi-Instance Editing

DGX agent

arXiv:2602.08749v3 Announce Type: replace Abstract: Flow matching models have recently emerged as an efficient alternative to diffusion, especially for text-guided image generation and editing, offeri

model-releasesarxiv-cs-cv
4 Jun 2026
Model Releases

Signed Dual Attention: Capturing Signed Dependencies in Time Series Forecasting

DGX agent

arXiv:2606.04833v1 Announce Type: cross Abstract: Initially developed for natural language processing, Transformer architectures and attention mechanisms are now central to a wide range of deep learni

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

SMAC-Talk: A Natural Language Extension of the StarCraft Multi-Agent Challenge for Large Language Models

DGX agent

arXiv:2606.04202v1 Announce Type: new Abstract: As LLMs become more widely deployed, they are increasingly expected to work alongside other AI agents rather than operating in isolation. Effective coor

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

SMADE-IE: Sparse Multi-Agent Framework with Evidence-Driven Debate for Zero-Shot Information Extraction

DGX agent

arXiv:2606.04691v1 Announce Type: new Abstract: Zero-shot information extraction (IE) with large language models (LLMs) has attracted increasing attention due to its flexibility in adapting to new sch

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

Solving Zebra Puzzles Using Constraint-Guided Multi-Agent Systems

DGX agent

arXiv:2407.03956v3 Announce Type: replace-cross Abstract: Prior research has enhanced the ability of Large Language Models (LLMs) to solve logic puzzles using techniques such as chain-of-thought promp

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

Sources: Anthropic has embedded around half a dozen forward-deployed engineers within the NSA to help the agency deploy Mythos for offensive cyber operations (Financial Times)

DGX agent

Financial Times: Sources: Anthropic has embedded around half a dozen forward-deployed engineers within the NSA to help the agency deploy Mythos for offensive cyber operations — Arrangement comes as AI

model-releasestechmeme
4 Jun 2026
Model Releases

Speculative Thinking: Enhancing Small-Model Reasoning with Large Model Guidance at Inference Time

DGX agent

arXiv:2504.12329v2 Announce Type: replace-cross Abstract: Recent advances leverage post-training to enhance model reasoning performance, which typically requires costly training pipelines and still su

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

SPLIT-PINN: Separable Probability Learning Technique via Physics-Informed Neural Networks for High-Dimensional Probabilistic Modeling

DGX agent

arXiv:2606.04000v1 Announce Type: cross Abstract: We present a probabilistic modeling framework for incorporating small-scale spatial heterogeneity into macroscopic descriptions of material behavior f

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

StandardE2E: A Unified Framework for End-to-End Autonomous Driving Datasets

DGX agent

arXiv:2606.04271v1 Announce Type: cross Abstract: Autonomous driving has shifted from modular perception-prediction-planning stacks toward end-to-end (E2E) models that map sensor inputs directly to ve

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

StepPRM-RTL: Stepwise Process-Reward Guided LLM Fine-Tuning for Enhanced RTL Synthesis

DGX agent

arXiv:2606.04246v1 Announce Type: new Abstract: Automatic generation of RTL code for digital hardware designs remains challenging due to long-horizon reasoning, multi-step dependencies, and strict cor

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Stepwise Reasoning Enhancement for LLMs via External Subgraph Generation

DGX agent

arXiv:2606.04454v1 Announce Type: new Abstract: Large language models have shown strong performance in natural language generation and downstream reasoning tasks, but they still struggle with logical

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

Streaming Communication in Multi-Agent Reasoning

DGX agent

arXiv:2606.05158v1 Announce Type: cross Abstract: Multi-agent reasoning systems adopt a 'generate-then-transfer' paradigm that forces end-to-end latency to scale linearly with pipeline depth. We intro

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

STRIDE: Training Data Attribution via Sparse Recovery from Subset Perturbations

DGX agent

arXiv:2606.05165v1 Announce Type: cross Abstract: Training Data Attribution (TDA) seeks to trace a model's predictions back to its training data. The gold standard for TDA relies on causal interventio

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

Structure-Aware Prediction of PROTAC-Mediated Protein Degradability via Graph Neural Networks

DGX agent

arXiv:2606.04021v1 Announce Type: cross Abstract: Proteolysis-targeting chimeras (PROTACs) can selectively degrade disease-causing proteins, yet predicting which targets are amenable to degradation re

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

Symbolic Regression for Shared Expressions: Introducing Partial Parameter Sharing

DGX agent

arXiv:2601.04051v3 Announce Type: replace Abstract: Symbolic regression aims to find symbolic expressions that describe datasets. Due to its inherent interpretability, symbolic regression (SR) is a po

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

SymTRELLIS: Symmetry-Enforced Voxel Latents for 3D Generation

DGX agent

arXiv:2606.04108v1 Announce Type: cross Abstract: Single-view 3D generative models have achieved impressive visual quality, yet they are not designed to satisfy structural or functional requirements,

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

TaDA: Calibrated Probe Gating for Task-Domain LoRA Merging

DGX agent

arXiv:2606.05016v1 Announce Type: new Abstract: Combining a task LoRA adapter with a domain LoRA adapter into a single unified model is a practical yet largely unexplored challenge. Existing methods t

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

Test-Time Compute Scaling for ASR with Depth-Conditioned Looped Transformers

DGX agent

arXiv:2606.04678v1 Announce Type: new Abstract: End-to-end ASR systems typically use fixed-depth acoustic encoders at inference, making it difficult to trade additional test-time computation for impro

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

That's a badass title, and it's true! Every day, it gets harder and harder to create tests that AI models can't beat. Reality is humanity's …

DGX agent

That's a badass title, and it's true! Every day, it gets harder and harder to create tests that AI models can't beat. Reality is humanity's real last exam. Andon Labs' Real-World AI Evals: Claude call

model-releasesswyx--x
4 Jun 2026
Model Releases

The Canadian AI strategy unveiled today advocates for the development of technology that is safe, ethical, trustworthy, and that benefits so…

DGX agent

The Canadian AI strategy unveiled today advocates for the development of technology that is safe, ethical, trustworthy, and that benefits society as a whole—these are exactly the principles that need

model-releasesyoshua-bengio--x
4 Jun 2026
Model Releases

The capabilities of Claude Code and Codex have expanded a lot in recent months, they added many ways to approach work (subagents, skills, go…

DGX agent

The capabilities of Claude Code and Codex have expanded a lot in recent months, they added many ways to approach work (subagents, skills, goal, workflows, plugins, etc). Given the AI labs can use thei

model-releasesethan-mollick--x
4 Jun 2026
← Previous
1…218219220221222…475
Next →