AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,563 results
Model Releases

Scaling Closed-Loop Feature Channel Configuration with LLMs

DGX agent

arXiv:2607.20516v1 Announce Type: cross Abstract: Promising initial results in closed-loop large-language-model-based channel-configuration search demonstrated that neural-network widths can be optimi

model-releasesarxiv-cs-ai
24 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Scaling Interpretable Transformers with Parity Bottleneck Layers

DGX agent

arXiv:2607.20652v1 Announce Type: cross Abstract: Language models are thought to exhibit the phenomenon of superposition, representing many more features than dimensions in their residual streams. Spa

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Scene Parameter Saliency via Differentiable Light Transport

DGX agent

arXiv:2607.21562v1 Announce Type: new Abstract: Gradient-based saliency methods reveal which input features most influence a neural network's output, and are a standard tool for model interpretability

model-releasesarxiv-cs-cv
24 Jul 2026
Model Releases

SciExplore: Evaluating Autonomous Agents from Scientific Navigation to Information Integration

DGX agent

arXiv:2607.20926v1 Announce Type: new Abstract: Scientific research involves complex information-seeking and reasoning workflows across heterogeneous sources. However, existing benchmarks primarily em

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Self-Evolving Recommendation System: End-To-End Autonomous Model Optimization With LLM Agents

DGX agent

arXiv:2602.10226v2 Announce Type: replace-cross Abstract: Optimizing large-scale machine learning systems, such as recommendation models for global video platforms, requires navigating a massive hyper

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Semi-Supervised Text-Attributed Graph Distillation

DGX agent

arXiv:2607.20477v1 Announce Type: new Abstract: {em Text-Attributed Graphs} (TAGs) have emerged as an expressive data model for integrating graph topology with rich textual semantics. Existing represe

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

SenCos-GEM: SENet-Calibrated and Law-of-Cosines-Constrained Geometry-Enhanced Molecular Representation for Property Prediction

DGX agent

arXiv:2607.20551v1 Announce Type: cross Abstract: Effective molecular representation learning is crucial for accurate molecular property prediction. Recently, numerous self-supervised learning (SSL) a

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

SESaMo: Symmetry-Enforcing Stochastic Modulation for Normalizing Flows

DGX agent

arXiv:2505.19619v3 Announce Type: replace Abstract: Deep generative models have recently garnered significant attention across various fields, from physics to chemistry, where sampling from unnormaliz

model-releasesarxiv-cs-lg
24 Jul 2026
Model Releases

Show, Don't Tell: Evaluating Spatial Cognition in Generative Pixels Rather Than LLM Text

DGX agent

arXiv:2607.21072v1 Announce Type: new Abstract: Spatial intelligence is essential for agents to move from static semantic understanding toward interacting with the physical world. Many spatial tasks a

model-releasesarxiv-cs-cv
24 Jul 2026
Model Releases

Silent Failures in Quantized LLM Reasoning: A Taxonomy-Based Analysis of Hollow Convergence and Failure Mode Shifts

DGX agent

arXiv:2607.09999v2 Announce Type: replace Abstract: We show that post-training quantization can silently alter how large language models reason even when task accuracy is preserved. Using a six-catego

model-releasesarxiv-cs-cl
24 Jul 2026
Model Releases

SkillCorpus: Consolidating and Evaluating the Open Skill Ecosystem for Real-World LLM Agents

DGX agent

arXiv:2607.15557v4 Announce Type: replace Abstract: Agent skills, SKILL files that package reusable procedural knowledge for an LLM agent, are a popular mechanism for extending agent capabilities. Pub

model-releasesarxiv-cs-cl
24 Jul 2026
Model Releases

SOAP, Muon, and Beyond: Pushing LLM Pretraining Scales

DGX agent

arXiv:2607.20548v1 Announce Type: cross Abstract: Higher-order optimizers such as Muon and SOAP offer faster convergence than AdamW, but their computational cost and numerical stability challenges hav

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

SonicSampler: Unified Tile-Aware Kernels for LLM Sampling and Speculative Verification

DGX agent

arXiv:2607.20475v1 Announce Type: new Abstract: Sampling in LLM inference comprises a combinatorial set of logit processing, token selection, and verification operations for speculative decoding. Howe

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Source-Prior-Driven Selective Adaptation for Efficient Diffusion Model Finetuning

DGX agent

arXiv:2607.20913v1 Announce Type: new Abstract: Fine-tuning large diffusion models for new domains or styles involves a trade-off: improving target-specific generation often degrades the pretrained mo

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Spectral functions in Minkowski quantum electrodynamics from neural reconstruction

DGX agent

arXiv:2510.24728v2 Announce Type: replace-cross Abstract: We study neural reconstructions of quenched rainbow quantum electrodynamics (QED) Dyson--Schwinger benchmarks in Minkowski-related kinematics.

model-releasesarxiv-cs-lg
24 Jul 2026
Model Releases

Spectral-Spatial Synergistic Guided Network for Hyperspectral Salient Object Detection

DGX agent

arXiv:2607.21032v1 Announce Type: new Abstract: Hyperspectral salient object detection aims to identify visually salient regions from hyperspectral images. Existing methods often fail because they fun

model-releasesarxiv-cs-cv
24 Jul 2026
Model Releases

StabilityBench: Benchmarking Instability in LLMs

DGX agent

arXiv:2607.20558v1 Announce Type: cross Abstract: AI Assistants are increasingly deployed in high-stakes settings, such as healthcare or government services. Yet their real-world behavior remains poor

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Statistical Inference for Generative Model Comparison

DGX agent

arXiv:2501.18897v4 Announce Type: replace-cross Abstract: Generative models have achieved remarkable success across a range of applications, yet their evaluation still lacks principled uncertainty qua

model-releasesarxiv-cs-lg
24 Jul 2026
Model Releases

Student-Centered Distillation Narrows the Agentic Gap Between Small and Large LLMs

DGX agent

arXiv:2509.14257v3 Announce Type: replace-cross Abstract: Large Language Model agents achieve strong performance on multi-step reasoning and tool-use tasks, but their impressive capabilities typically

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

swiss-ai/Apertus-v1.5 70B/8B

DGX agent

https://huggingface.co/swiss-ai/Apertus-v1.5-70B https://huggingface.co/swiss-ai/Apertus-v1.5-8B Apertus 1.5 is a family of 8B and 70B parameter language models designed to advance the state of multil

model-releasesr-localllama
24 Jul 2026
Model Releases

T-STAR: A Large-Scale Benchmark for Spatio-Temporal Panoptic Scene Graph Generation in Satellite Video

DGX agent

arXiv:2607.21228v1 Announce Type: new Abstract: Structured understanding of satellite video is essential for advancing dynamic geospatial scene analysis from low-level perception to high-level cogniti

model-releasesarxiv-cs-cv
24 Jul 2026
Model Releases

Telco-GAIA: Bilingual Benchmark for Agents in Telecom Domain

DGX agent

arXiv:2607.20510v1 Announce Type: new Abstract: We introduce Telco-GAIA, a bilingual, multi-modal benchmark for evaluating tool-using agents on the data of a real-world telecommunications operator. Te

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Tencent WorkBuddy Bench: A Multi-Domain Coding-Agent Benchmark with Contamination-Resistant Task Construction

DGX agent

arXiv:2607.20911v1 Announce Type: new Abstract: We introduce Tencent WorkBuddy Bench, a multi-domain evaluation suite for coding agents; this report documents its construction methodology, scoring pro

model-releasesarxiv-cs-cl
24 Jul 2026
Model Releases

The Devil is in the Spectrum: Mitigating Representation Collapse in LLMs via Topologically Regularized Side-Path

DGX agent

arXiv:2607.20484v1 Announce Type: new Abstract: Large Language Models (LLMs) are fundamentally limited by representation collapse, a bottleneck that severely degrades long-context performance. We iden

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

The Geometry of Personality: Activation Steering with Jungian Cognitive Functions

DGX agent

arXiv:2607.20803v1 Announce Type: cross Abstract: Activation steering enables control and interpretation of LLMs, yet existing work primarily models personality through static trait frameworks such as

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

The GPT-5.x Pro series has remained the best models for hard technical problems since they launched. There is some parallel model magic goin…

DGX agent

The GPT-5.x Pro series has remained the best models for hard technical problems since they launched. There is some parallel model magic going on that is not well-explained. Anthropic has never had an

model-releasesethan-mollick--x
24 Jul 2026
Model Releases

The Hidden Footprint: Making Storage a First-Class Metric for LLM Agent Evaluation

DGX agent

arXiv:2607.11149v3 Announce Type: replace Abstract: LLM agent benchmarks measure task completion, reliability, and inference cost, but not the persistent data an agent run leaves on disk, including lo

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

The RealDefocus Benchmark for Defocus Deblurring

DGX agent

arXiv:2607.21078v1 Announce Type: new Abstract: Single-Image Defocus Deblurring (SIDD) aims to recover an all-in-focus image from a single defocused observation, but rigorous and reproducible evaluati

model-releasesarxiv-cs-cv
24 Jul 2026
Model Releases

The Second LoViF 2026 Challenge on Real-World All-in-One Image Restoration: Methods and Results

DGX agent

arXiv:2607.21118v1 Announce Type: new Abstract: This paper presents a review of the second LoViF Challenge on Real-World All-in-One Image Restoration. The challenge aims to advance unified image resto

model-releasesarxiv-cs-cv
24 Jul 2026
Model Releases

The Washington Post processed 1.79B input tokens per month through Together AI, running open models like Llama and Mistral in production wit…

DGX agent

The Washington Post processed 1.79B input tokens per month through Together AI, running open models like Llama and Mistral in production with predictable costs and full control over the model stack. T

model-releasestogether-ai--x
24 Jul 2026
Model Releases

This is a big jump in ARC-AGI-3.

DGX agent

This is a big jump in ARC-AGI-3. Claude Opus 5 from @AnthropicAI is the new SOTA on ARC-AGI-3: 30.2% The previous high score (7.8%) was set by GPT-5.6 Sol (Max) Throughout our analysis, we observed no

model-releasesethan-mollick--x
24 Jul 2026
Model Releases

Three-Pronged Spectral Control for Federated Parameter Efficient Fine Tuning

DGX agent

arXiv:2607.20914v1 Announce Type: new Abstract: Federated parameter-efficient fine-tuning (PEFT) enables communication-efficient adaptation of large pretrained models on decentralized edge data, but i

model-releasesarxiv-cs-lg
24 Jul 2026
Model Releases

Token Budget Saturation and Mechanistic Early Detection of Reasoning Non-Convergence in Chain-of-Thought Models

DGX agent

arXiv:2607.21433v1 Announce Type: cross Abstract: Chain-of-thought reasoning models such as DeepSeek-R1-Distill-Qwen-7B exhibit a bimodal convergence pattern: generations either terminate within a tok

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

TOPReward: Token Probabilities as Hidden Zero-Shot Rewards for Robotics

DGX agent

arXiv:2602.19313v2 Announce Type: replace-cross Abstract: General-purpose robot learning requires dense, instruction-conditioned feedback that can distinguish meaningful task progress from stalled, fa

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

torchsom: The Reference PyTorch Library for Self-Organizing Maps

DGX agent

arXiv:2510.11147v2 Announce Type: replace-cross Abstract: This paper introduces torchsom, an open-source Python library that provides a reference implementation of the Self-Organizing Map (SOM) in PyT

model-releasesarxiv-cs-lg
24 Jul 2026
Model Releases

TOUR: A Trajectory-Level Unlearning Benchmark for Offline Reinforcement Learning

DGX agent

arXiv:2607.21111v1 Announce Type: cross Abstract: Offline Reinforcement Learning (RL) agents are trained on fixed behavioral trajectories, which makes trajectory-level deletion important when selected

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Toward Generalizable Cognitive Impairment Detection with Speech-Based Multimodal Large Language Models

DGX agent

arXiv:2607.21496v1 Announce Type: cross Abstract: Cognitive impairment (CI) is a growing public health concern. Early and accurate diagnosis is critical for enabling timely intervention and improving

model-releasesarxiv-cs-lg
24 Jul 2026
Model Releases

Toward Mechanistic Interpretability of an AI Foundation Model Fine-Tuned for Atmospheric Chemistry

DGX agent

arXiv:2607.20778v1 Announce Type: new Abstract: Weather forecasting foundation models (FMs) are increasingly fine-tuned to predict air quality, offering fast global pollution forecasts at lower comput

model-releasesarxiv-cs-lg
24 Jul 2026
Model Releases

Towards an Automated Test of LLM Security Knowledge

DGX agent

arXiv:2607.18496v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used for a range of software, hardware and human-centered security tasks. Consequently, LLM perf

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Towards Faithful Graph Explanations with Synergistic Edge Effects via Granular Balls

DGX agent

arXiv:2607.21381v1 Announce Type: new Abstract: Instance-level explanations aim to reveal the rationale behind a model's decisions for a specific graph. Previous methods explain graph neural networks

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Training Large Language Models for Self-Explanation Faithfulness

DGX agent

arXiv:2607.21090v1 Announce Type: cross Abstract: We propose a Reinforcement Learning (RL) method to directly optimize the faithfulness of self-explanations - the extent to which a model's generated r

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

TransBiolab: A Real-World Multi-View Dataset of Cluttered Transparent Biomedical Objects

DGX agent

arXiv:2607.21071v1 Announce Type: new Abstract: Autonomous biomedical laboratories increasingly rely on visual perception to recognize, localize, and manipulate transparent plasticware, yet high-quali

model-releasesarxiv-cs-cv
24 Jul 2026
Model Releases

U-CFR: Uncertainty-Guided Cascade Forward Refinement for Interactive Segmentation

DGX agent

arXiv:2607.20705v1 Announce Type: cross Abstract: Interactive image segmentation is critical for efficient image annotation; however, existing methods often require many corrective clicks or rely on p

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Unlearning Under Imbalance: Benchmarking Fairness in Multimodal LLM Unlearning

DGX agent

arXiv:2607.21300v1 Announce Type: cross Abstract: Machine unlearning has emerged as a tool for removing personal data from trained models to comply with recent AI regulations. To evaluate unlearning e

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Verifier-First Evaluation of Agentic LLMs for Infrastructure-as-Code Generation

DGX agent

arXiv:2607.20478v1 Announce Type: cross Abstract: Infrastructure-as-Code (IaC) generation from natural language requires satisfying provider schemas, dependency planning, and organizational policy con

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

ViSTR-Bench: Can MLLMs Reason from Continuous Visual Cues in Dynamic Scenes?

DGX agent

arXiv:2607.20868v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable success across diverse expert-level tasks, but they still struggle with fundamental ab

model-releasesarxiv-cs-cv
24 Jul 2026
Model Releases

Visual Contrastive Self-Distillation

DGX agent

arXiv:2607.21556v1 Announce Type: cross Abstract: On-policy self-distillation (OPSD) is promising as it removes the external teacher required by on-policy distillation (OPD), yet it still needs asymme

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

VoLN: Vision-Only Long-Horizon Navigation---Paradigm, Benchmark, and Method

DGX agent

arXiv:2607.21400v1 Announce Type: cross Abstract: Vision-and-Language Navigation (VLN) enables embodied agents to follow natural-language instructions. However, route-level instructions commonly encod

model-releasesarxiv-cs-ai
24 Jul 2026
← Previous
1…9192939495…471
Next →