AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
86,428Total entries
1Added by human
86,427Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
23,185 results
Model Releases

Struct-Searcher: Agentic Structural Thinking Advances Multimodal Deep Information Seeking

DGX agent

arXiv:2606.07689v1 Announce Type: new Abstract: Deep research agents have attracted increasing attention for their ability to collect large-scale online information to acquire target knowledge, with r

model-releasesarxiv-cs-cv
9 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Structured Neuron Pruning in Deep Neural Networks Using Multi-Armed Bandits

DGX agent

arXiv:2606.07615v1 Announce Type: cross Abstract: Deep neural networks often contain redundant hidden units. Removing individual weights can reduce parameter count, but unstructured sparsity is not al

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Students without access to LLMs are 2 to 8 times more creative than students with access. That is the finding of a new paper comparing 2,200…

DGX agent

Students without access to LLMs are 2 to 8 times more creative than students with access. That is the finding of a new paper comparing 2,200 college admissions essays written by humans before ChatGPT

model-releasesgary-marcus--x
9 Jun 2026
Model Releases

Subtitle-Aligned Fine-Tuning of Whisper for Swiss German ASR: Benchmark Contamination, Convention Mismatch, and an Honest Baseline at 25.6% WER (13.8% cWER)

DGX agent

arXiv:2606.07608v1 Announce Type: cross Abstract: We present a systematic study of fine-tuning OpenAI's Whisper large-v3 for Swiss German ASR, using 1,367 hours of broadcast speech paired with Standar

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Supracompetitive Pricing Under AI Monoculture

DGX agent

arXiv:2601.01279v3 Announce Type: replace-cross Abstract: When competing sellers delegate pricing to a shared AI model, such as a large language model, correlated recommendations combined with perform

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

SurfDesign: Effective Protein Design on Molecular Surfaces

DGX agent

arXiv:2606.07567v1 Announce Type: cross Abstract: Protein function is largely determined by molecular surface geometry and physicochemical complementarity, yet most protein design methods condition on

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

SWE-Marathon: Can Agents Autonomously Complete Ultra-Long-Horizon Software Work?

DGX agent

arXiv:2606.07682v1 Announce Type: cross Abstract: AI agents are increasingly expected to complete long-horizon workflows that require sustained progress over hours, millions of tokens, and complex env

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Systematic LLM Translation of Legacy Scientific Code to Differentiable Frameworks: Application to a Land Surface Model

DGX agent

arXiv:2606.07681v1 Announce Type: cross Abstract: Differentiable programming offers transformative capabilities for scientific modeling, enabling gradient-based parameter estimation, sensitivity analy

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

TABVERSE: Benchmarking Cross-Format Table Understanding in LLMs and VLMs

DGX agent

arXiv:2606.09578v1 Announce Type: new Abstract: Large Language Models (LLMs) and Vision-Language Models (VLMs) are increasingly evaluated on table reasoning tasks, but the role of table representation

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

TAME: A Trustworthy Test-Time Evolution of Agent Memory with Systematic Benchmarking

DGX agent

arXiv:2602.03224v2 Announce Type: replace Abstract: Test-time evolution of agent memory represents a pivotal paradigm for advancing AGI, as it strengthens complex reasoning through experience accumula

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

TempoBench: Evaluating Temporal Causal Reasoning in Large Language Models

DGX agent

arXiv:2510.27544v2 Announce Type: replace Abstract: Temporal reasoning involves understanding how systems evolve over time through input-driven state transitions. A key aspect is temporal causal reaso

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

The AI Epistemic Deference Index: A Continuous Measure of Sycophancy

DGX agent

arXiv:2606.07897v1 Announce Type: new Abstract: Current AI models frequently exhibit epistemic sycophancy, endorsing claims to agree with a user. Existing evaluations typically measure this either by

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

The Injection Paradox: Brand-Level Suppression in Safety-Trained LLM Recommendations via RAG Context Injection

DGX agent

arXiv:2606.09204v1 Announce Type: new Abstract: We present a reproducible failure mode of safety training in RAG-based LLM recommendation -- the Injection Paradox -- in which prompt injections embedde

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

The Last Visible Pixel: Probing Fine-Scale Perception in Vision-Language Models

DGX agent

arXiv:2606.07861v1 Announce Type: cross Abstract: Recent vision-language models (VLMs) excel at multimodal understanding and reasoning, yet their fine-grained visual perception remains underexplored.

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

The Montparnasse Algorithm for RNA Design

DGX agent

arXiv:2606.07562v1 Announce Type: cross Abstract: RNA design consists of discovering a nucleotide sequence that optimizes predefined criteria, such as secondary structure. It is useful for synthetic b

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

The Sample Complexity of Parameter-Free Stochastic Convex Optimization

DGX agent

arXiv:2506.11336v2 Announce Type: replace Abstract: We study the sample complexity of stochastic convex optimization when problem parameters such as the distance to optimality and the Lipschitz consta

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

TheoremBench: Evaluating LLMs on Theorem Proving in Formal Mathematics

DGX agent

arXiv:2606.09450v1 Announce Type: new Abstract: LLMs have recently achieved strong results on formal proving benchmarks. However, existing evaluations remain heavily concentrated on competition-style

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Theoretical Foundations of Continual Learning via Drift-Plus-Penalty

DGX agent

arXiv:2606.08452v1 Announce Type: new Abstract: In many real-world settings, data streams are nonstationary and arrive sequentially, requiring learning systems to adapt continuously without retraining

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

There has been a lot of hand wringing on the appropriate valuation of SpaceX. Some large institutions believe SpaceX can only be valued at h…

DGX agent

There has been a lot of hand wringing on the appropriate valuation of SpaceX. Some large institutions believe SpaceX can only be valued at half what the market seems to be willing to pay for it. Other

model-releaseselon-musk--x
9 Jun 2026
Model Releases

They ruled Iryna’s killer is incompetent to stand trial. The same system ruled this man was plenty competent enough to be released back into…

DGX agent

They ruled Iryna’s killer is incompetent to stand trial. The same system ruled this man was plenty competent enough to be released back into society dozens of times. It’s past time to remove these lef

model-releaseselon-musk--x
9 Jun 2026
Model Releases

Thinking-Based Non-Thinking: Solving the Reward Hacking Problem in Training Hybrid Reasoning Models via Reinforcement Learning

DGX agent

arXiv:2601.04805v2 Announce Type: replace Abstract: Large reasoning models (LRMs) have attracted much attention due to their exceptional performance. However, their performance mainly stems from think

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

This is a super exciting release - Claude Fable 5 is the same underlying model as Mythos but with added safeguards. The benchmarks are great…

DGX agent

This is a super exciting release - Claude Fable 5 is the same underlying model as Mythos but with added safeguards. The benchmarks are great and it's SOTA on everything by a margin but I'll add that *

model-releaseskarpathy--x
9 Jun 2026
Model Releases

Tiger Data launches PostgreSQL extension designed for AI agents

DGX agent

Tiger Data today introduced a managed PostgreSQL database service designed specifically for AI agents, saying conventional database architectures are poorly suited to a future in which software is inc

model-releasessiliconangle
9 Jun 2026
Model Releases

Today, we released Gemini 3.5 Live Translate, our latest audio model for live speech-to-speech translation. It supports over 70 languages an…

DGX agent

Today, we released Gemini 3.5 Live Translate, our latest audio model for live speech-to-speech translation. It supports over 70 languages and starts translating as soon as you start talking, streaming

model-releasesgoogle-ai--x
9 Jun 2026
Model Releases

Token Sample Complexity of Attention

DGX agent

arXiv:2512.10656v3 Announce Type: replace Abstract: As context windows in large language models continue to expand, it is essential to characterize how attention behaves at extreme sequence lengths. W

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Towards Long-Horizon Vessel Trajectory and Destination Forecasting with Reasoning Large Language Models

DGX agent

arXiv:2606.08633v1 Announce Type: new Abstract: Long-horizon maritime trajectory prediction is important for shipping management, logistics planning, and maritime risk analysis, yet month-level foreca

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Towards Personalized Bangla Book Recommendation: A Large-Scale Heterogeneous Book Graph Dataset

DGX agent

arXiv:2602.12129v2 Announce Type: replace-cross Abstract: Personalized book recommendation in Bangla literature has been constrained by the lack of structured, large-scale, and publicly available data

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

TQA-Bench: Evaluating LLMs for Multi-Table Question Answering

DGX agent

arXiv:2411.19504v2 Announce Type: replace Abstract: The advance of large language models (LLMs) has unlocked great opportunities in complex multi-modal data management tasks, particularly in question

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Training-Free Generalized Few-Shot Segmentation through Open-Vocabulary Semantic Arbitration

DGX agent

arXiv:2606.09474v1 Announce Type: new Abstract: Generalized Few-Shot Semantic Segmentation (GFSS) has traditionally been approached as a representation-learning problem, requiring task-specific adapta

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Trajectory-Refined Distillation

DGX agent

arXiv:2606.08432v1 Announce Type: new Abstract: On-policy distillation (OPD) has become a central post-training tool for large language models (LLMs), providing dense per-token teacher supervision alo

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

TriHead-GAN: A Generative Adversarial Network with Triple-Head Discriminator for Carbon Emission Time Series Generation

DGX agent

arXiv:2606.07569v1 Announce Type: new Abstract: Accurate carbon emission monitoring is critical for climate policy and emerging regulatory mechanisms such as the EU Carbon Border Adjustment Mechanism,

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

TRL-Bench: Standardizing Cross-Paradigm Representation-Level Evaluation of Tabular Encoders

DGX agent

arXiv:2606.09323v1 Announce Type: new Abstract: Tabular encoders are usually evaluated inside task-specific end-to-end pipelines, so models from different training paradigms are difficult to compare d

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

TT-DAC-PS: Twin-Target Deterministic Actor-Critic with Policy Smoothing for Optimal Trade Execution

DGX agent

arXiv:2606.08379v1 Announce Type: new Abstract: This study addresses the optimal execution of large stock sell programs by introducing TT-DAC-PS (Twin-Target Deterministic Actor-Critic with Policy Smo

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Understanding Benchmark Language Under Weakened Formal Semantics

DGX agent

arXiv:2509.17455v2 Announce Type: replace-cross Abstract: State-of-the-art NLP benchmarks require interpretation of natural language that specifies conditions, procedures, and exceptions, often relyin

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Understanding the Parameter Space Geometry of Transformers Encoding Boolean Functions

DGX agent

arXiv:2606.08768v1 Announce Type: new Abstract: Transformers consistently fail to learn certain simple functions that are provably expressible with specific parameter settings. This gap between learna

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Unification of Closed-Open Industrial Detection Scenarios: New Large-Scale Benchmarks,Challenges and Baselines

DGX agent

arXiv:2606.07953v1 Announce Type: new Abstract: Large-scale Visual-Language Models (LVLMs) have achieved remarkable success in natural visual tasks, yet their application to industrial defect detectio

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

UniQL: Towards Dialect-Universal Benchmarking for Text-to-SQL

DGX agent

arXiv:2606.08018v1 Announce Type: new Abstract: Existing text-to-SQL benchmarks are largely centered on SQLite, making it difficult to evaluate whether models can generalize across heterogeneous SQL d

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Unsupervised Partner Design Enables Robust Ad-hoc Teamwork

DGX agent

arXiv:2508.06336v2 Announce Type: replace-cross Abstract: We introduce Unsupervised Partner Design (UPD), a population-free multi-agent reinforcement learning method for robust ad-hoc teamwork. UPD ge

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

VATS: Exploiting Implicit Authority in Error-Path Injection via Systematic Mutation

DGX agent

arXiv:2606.07992v1 Announce Type: new Abstract: As the Model Context Protocol (MCP) standardizes tool-calling for autonomous agents, it introduces a critical, unexamined attack surface: the error-hand

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Version of AI tool too powerful for public released to public https://bbc.in/4xfGSlq

DGX agent

A version of an AI tool deemed too powerful for public release was inadvertently made available to the public, according to a report from BBC News. The incident highlights concerns about controlling a

model-releasesclem-delangue--x
9 Jun 2026
Model Releases

VideoWeaver: Evaluating and Evolving Skills for Agentic Long Video Generation

DGX agent

arXiv:2606.08091v1 Announce Type: new Abstract: Recent agent frameworks such as Claude Code, Codex, and OpenClaw are strong at tool use and orchestration, but whether they can handle long video genera

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Visual Template Inference for Data Extraction from Documents

DGX agent

arXiv:2501.06659v2 Announce Type: replace-cross Abstract: Many templatized documents are programmatically generated from structured data following a visual template. Such documents include invoices, t

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning?

DGX agent

arXiv:2606.07872v1 Announce Type: new Abstract: When a multimodal large language model answers a visual reasoning question correctly, is the prediction actually supported by the task-critical visual e

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

VisualLeakBench: Reproducible Action-Boundary Propagation Failures in Vision-Language Agents

DGX agent

arXiv:2606.07595v1 Announce Type: cross Abstract: Vision-language agents increasingly consume screenshots, documents, and user interfaces before writing to memory, sending messages, or invoking extern

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

vla.cpp: A Unified Inference Runtime for Vision-Language-Action Models

DGX agent

arXiv:2606.08094v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) policies are typically shipped as Python/PyTorch stacks that assume a workstation-class GPU, a mismatch for the hardware

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

VoLo: A Physical Orchestrator for Open-Vocabulary Long-Horizon Manipulation

DGX agent

arXiv:2606.07723v1 Announce Type: new Abstract: Open-vocabulary long-horizon manipulation requires robots to reason over flexible instructions and complex multi-object scenes while adaptively planning

model-releasesarxiv-cs-ro
9 Jun 2026
Model Releases

We encourage developers to share their builds with us and give feedback to shape future iterations. Let’s shape the future of sovereign AI t…

DGX agent

We encourage developers to share their builds with us and give feedback to shape future iterations. Let’s shape the future of sovereign AI together. Download: https://huggingface.co/CohereLabs/North-M

model-releasescohere--x
9 Jun 2026
Model Releases

We talk a lot about how important it is to set up self-verification loops. Especially in the age of powerful models that can run for long pe…

DGX agent

We talk a lot about how important it is to set up self-verification loops. Especially in the age of powerful models that can run for long periods of time, self-verification is a key ingredient that en

model-releasesboris-cherny--x
9 Jun 2026
← Previous
1…212213214215216…484
Next →