AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

Scaling Laws for Masked-Reconstruction Transformers on Single-Cell Transcriptomics

DGX agent

arXiv:2602.15253v2 Announce Type: replace Abstract: Neural scaling laws -- power-law relationships between loss, model size, and data -- have been extensively documented for language and vision transf

model-releasesarxiv-cs-lg
9 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

scCBGM: Interpretable Single-Cell Counterfactual Editing

DGX agent

arXiv:2606.07760v1 Announce Type: new Abstract: Understanding cellular phenotypes and how they respond to perturbations is critical for disease biology and therapeutic design. Single-cell RNA sequenci

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

SceneConductor: 3D Scene Generation from Single Image with Multi-Agent Orchestration

DGX agent

arXiv:2606.08402v1 Announce Type: cross Abstract: Generating complete 3D scenes from a single image requires inferring globally consistent geometry, object relationships, and environmental context fro

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Sci-Rho: A Multilingual Visually-Grounded Symbolic Benchmark for STEM Problems

DGX agent

arXiv:2606.08034v1 Announce Type: cross Abstract: Symbolic benchmarks have emerged as a key approach to assess model robustness under minor modifications to STEM-related questions. However, existing s

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

SciFlow-Bench: Evaluating Structure-Aware Scientific Diagram Generation via Inverse Parsing

DGX agent

arXiv:2602.09809v2 Announce Type: replace Abstract: Scientific diagrams convey explicit structural information, yet modern text-to-image models often produce visually plausible but structurally incorr

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

See More, Think Deeper: Query-Expanded Visual Evidence and Answer-Clue Guided Reflection for Long Video Understanding

DGX agent

arXiv:2606.09064v1 Announce Type: cross Abstract: Recent advances in Video Large Language Models (Video-LLMs) have enabled performance on long-video understanding tasks. However, existing methods stil

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

SegmentAnyTreeV2: Scaling Transformer-Based Tree Instance Segmentation Across Sensors, Platforms, and Forests

DGX agent

arXiv:2606.08206v1 Announce Type: new Abstract: We present SegmentAnyTreeV2, a sensor- and platform-agnostic framework for semantic and instance segmentation of forest point clouds. The model combines

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Semi-supervised Source Detection in Astronomical Images: New Benchmark and Strong Baseline

DGX agent

arXiv:2606.09219v1 Announce Type: new Abstract: Source detection in modern observational astronomy is a cornerstone for localizing and identifying stellar sources accurately. It is crucial for studies

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

SENTRY: Statistical Reliability Analysis of Vision Transformers Under Soft Errors

DGX agent

arXiv:2606.07620v1 Announce Type: cross Abstract: With the growth of Vision Transformers in safety-critical domains like autonomous systems and medical imaging, ensuring their reliability against soft

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Seq103: A Unified Neuroevolution Framework for Compact Sequence Architecture Discovery

DGX agent

arXiv:2606.07664v1 Announce Type: cross Abstract: Neuroevolution is a representative neural architecture search paradigm that evolves both network topology and weights through evolutionary algorithms.

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Shared Latent Structures Enable Unified Backdoor Detection and Mitigation in LLMs

DGX agent

arXiv:2606.07963v1 Announce Type: new Abstract: Backdoor attacks in large language models (LLMs) are often treated as isolated trigger-response failures, motivating defenses tailored to specific trigg

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Shift-Dependent Asymmetry: Orthogonal Inverse Low-Rank Adaptation for Federated Medical Segmentation

DGX agent

arXiv:2606.08687v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) enables efficient federated fine-tuning of segmentation foundation models for medical imaging. However, most federated LoRA m

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Signals Are Not States: Neuro-Symbolic Safeguards for Culturally Aware Classroom AI

DGX agent

arXiv:2603.22793v2 Announce Type: replace Abstract: Classroom AI systems increasingly infer high-level educational states such as engagement, confusion, collaboration, participation, and instructional

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

SIMPLE: Simulation-Based Policy Learning and Evaluation for Humanoid Loco-manipulation

DGX agent

arXiv:2606.08278v1 Announce Type: new Abstract: Humanoid foundation models are advancing faster than we can evaluate them. While real-world testing is expensive and difficult to reproduce, existing si

model-releasesarxiv-cs-ro
9 Jun 2026
Model Releases

SLMJury: Can Small Language Models Judge as Well as Large Ones?

DGX agent

arXiv:2606.07810v1 Announce Type: cross Abstract: Large language models (LLMs) are widely used as judges for evaluating model outputs, but their high cost, latency, and opacity limit scalability. We i

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

SNN-MLIR: An MLIR Dialect for Compiling Neuromorphic SNNs from NIR to Bare-Metal C

DGX agent

arXiv:2606.09213v1 Announce Type: cross Abstract: Spiking neural networks (SNNs) are increasingly trained in a wide range of frameworks (SnnTorch, Lava, Norse, and others) each with its own model form

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

SoK: Reconstruction Attacks on Synthetic Tabular Data (Insights from Winning the NIST CRC)

DGX agent

arXiv:2606.08372v1 Announce Type: cross Abstract: Synthetic data is increasingly promoted as a privacy-preserving substitute for releasing sensitive tabular records, yet its central adversarial threat

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Solving Inverse Problems with Flow-based Models via Model Predictive Control

DGX agent

arXiv:2601.23231v2 Announce Type: replace-cross Abstract: Flow-based generative models provide strong unconditional priors for inverse problems, but guiding their dynamics for conditional generation r

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Sparse Autoencoders Reveal Interpretable and Steerable Features in VLA Models

DGX agent

arXiv:2603.19183v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models have emerged as a promising approach for general-purpose robot manipulation. However, little research has mechan

model-releasesarxiv-cs-ro
9 Jun 2026
Model Releases

SpatialWorld: Benchmarking Interactive Spatial Reasoning of Multimodal Agents in Real-World Tasks

DGX agent

arXiv:2606.09669v1 Announce Type: new Abstract: Spatial reasoning is a foundational capability for multimodal large language models (MLLMs) to perceive and operate within the physical world. However,

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

SpectrumKV: Per-Token Mixed-Precision KV Cache Transfer for Prefill-Decode Disaggregated LLM Serving

DGX agent

arXiv:2606.08635v1 Announce Type: new Abstract: Prefill-decode (PD) disaggregation decouples prompt processing from token generation, but it also turns the key-value (KV) cache into a network payload.

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

SSAFE: Simple and Strong AI-Generated Image Detection via Frozen Vision Encoders

DGX agent

arXiv:2606.08634v1 Announce Type: new Abstract: The rapid advancement of generative models has blurred the boundary between synthetic and real imagery, creating an urgent need for reliable deepfake de

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Stabilizing On-Policy Distillation for MLLM Reasoning with Global Normalization

DGX agent

arXiv:2606.09091v1 Announce Type: cross Abstract: On-policy distillation (OPD) has recently emerged as an important post-training paradigm. By using a stronger teacher model to provide dense, fine-gra

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Still: Amortized KV Cache Compaction in a Single Forward Pass

DGX agent

arXiv:2606.07878v1 Announce Type: new Abstract: The KV cache is the memory bottleneck of long-horizon language model deployment. Practically, a deployable compactor must be lightweight enough to call

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Strained Coherence: A Pre-Failure Signal in Coding Agent Execution Trajectories

DGX agent

arXiv:2606.07889v1 Announce Type: cross Abstract: LLM-based coding agents sometimes acknowledge a problem in their own reasoning and then proceed anyway. We call this pattern strained coherence: a saf

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Streaming Interventions: Can Video Large Language Models Correct Mistakes as They Occur?

DGX agent

arXiv:2606.09547v1 Announce Type: new Abstract: Learning everyday skills, like cooking a dish, relies increasingly on instructional media such as online videos. This opens the door to the use of video

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Stress-testing medical large language models reveals latent safety pathology beyond benchmark accuracy

DGX agent

arXiv:2606.07929v1 Announce Type: new Abstract: Large language models (LLMs) are entering clinical practice based on benchmark accuracy that may fail to detect safety-relevant failure modes. Here we p

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Struct-Searcher: Agentic Structural Thinking Advances Multimodal Deep Information Seeking

DGX agent

arXiv:2606.07689v1 Announce Type: new Abstract: Deep research agents have attracted increasing attention for their ability to collect large-scale online information to acquire target knowledge, with r

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Structured Neuron Pruning in Deep Neural Networks Using Multi-Armed Bandits

DGX agent

arXiv:2606.07615v1 Announce Type: cross Abstract: Deep neural networks often contain redundant hidden units. Removing individual weights can reduce parameter count, but unstructured sparsity is not al

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Subtitle-Aligned Fine-Tuning of Whisper for Swiss German ASR: Benchmark Contamination, Convention Mismatch, and an Honest Baseline at 25.6% WER (13.8% cWER)

DGX agent

arXiv:2606.07608v1 Announce Type: cross Abstract: We present a systematic study of fine-tuning OpenAI's Whisper large-v3 for Swiss German ASR, using 1,367 hours of broadcast speech paired with Standar

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Supracompetitive Pricing Under AI Monoculture

DGX agent

arXiv:2601.01279v3 Announce Type: replace-cross Abstract: When competing sellers delegate pricing to a shared AI model, such as a large language model, correlated recommendations combined with perform

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

SurfDesign: Effective Protein Design on Molecular Surfaces

DGX agent

arXiv:2606.07567v1 Announce Type: cross Abstract: Protein function is largely determined by molecular surface geometry and physicochemical complementarity, yet most protein design methods condition on

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

SWE-Marathon: Can Agents Autonomously Complete Ultra-Long-Horizon Software Work?

DGX agent

arXiv:2606.07682v1 Announce Type: cross Abstract: AI agents are increasingly expected to complete long-horizon workflows that require sustained progress over hours, millions of tokens, and complex env

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Systematic LLM Translation of Legacy Scientific Code to Differentiable Frameworks: Application to a Land Surface Model

DGX agent

arXiv:2606.07681v1 Announce Type: cross Abstract: Differentiable programming offers transformative capabilities for scientific modeling, enabling gradient-based parameter estimation, sensitivity analy

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

TABVERSE: Benchmarking Cross-Format Table Understanding in LLMs and VLMs

DGX agent

arXiv:2606.09578v1 Announce Type: new Abstract: Large Language Models (LLMs) and Vision-Language Models (VLMs) are increasingly evaluated on table reasoning tasks, but the role of table representation

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

TAME: A Trustworthy Test-Time Evolution of Agent Memory with Systematic Benchmarking

DGX agent

arXiv:2602.03224v2 Announce Type: replace Abstract: Test-time evolution of agent memory represents a pivotal paradigm for advancing AGI, as it strengthens complex reasoning through experience accumula

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

TempoBench: Evaluating Temporal Causal Reasoning in Large Language Models

DGX agent

arXiv:2510.27544v2 Announce Type: replace Abstract: Temporal reasoning involves understanding how systems evolve over time through input-driven state transitions. A key aspect is temporal causal reaso

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

The AI Epistemic Deference Index: A Continuous Measure of Sycophancy

DGX agent

arXiv:2606.07897v1 Announce Type: new Abstract: Current AI models frequently exhibit epistemic sycophancy, endorsing claims to agree with a user. Existing evaluations typically measure this either by

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

The Injection Paradox: Brand-Level Suppression in Safety-Trained LLM Recommendations via RAG Context Injection

DGX agent

arXiv:2606.09204v1 Announce Type: new Abstract: We present a reproducible failure mode of safety training in RAG-based LLM recommendation -- the Injection Paradox -- in which prompt injections embedde

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

The Last Visible Pixel: Probing Fine-Scale Perception in Vision-Language Models

DGX agent

arXiv:2606.07861v1 Announce Type: cross Abstract: Recent vision-language models (VLMs) excel at multimodal understanding and reasoning, yet their fine-grained visual perception remains underexplored.

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

The Montparnasse Algorithm for RNA Design

DGX agent

arXiv:2606.07562v1 Announce Type: cross Abstract: RNA design consists of discovering a nucleotide sequence that optimizes predefined criteria, such as secondary structure. It is useful for synthetic b

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

The Sample Complexity of Parameter-Free Stochastic Convex Optimization

DGX agent

arXiv:2506.11336v2 Announce Type: replace Abstract: We study the sample complexity of stochastic convex optimization when problem parameters such as the distance to optimality and the Lipschitz consta

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

TheoremBench: Evaluating LLMs on Theorem Proving in Formal Mathematics

DGX agent

arXiv:2606.09450v1 Announce Type: new Abstract: LLMs have recently achieved strong results on formal proving benchmarks. However, existing evaluations remain heavily concentrated on competition-style

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Theoretical Foundations of Continual Learning via Drift-Plus-Penalty

DGX agent

arXiv:2606.08452v1 Announce Type: new Abstract: In many real-world settings, data streams are nonstationary and arrive sequentially, requiring learning systems to adapt continuously without retraining

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Thinking-Based Non-Thinking: Solving the Reward Hacking Problem in Training Hybrid Reasoning Models via Reinforcement Learning

DGX agent

arXiv:2601.04805v2 Announce Type: replace Abstract: Large reasoning models (LRMs) have attracted much attention due to their exceptional performance. However, their performance mainly stems from think

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Token Sample Complexity of Attention

DGX agent

arXiv:2512.10656v3 Announce Type: replace Abstract: As context windows in large language models continue to expand, it is essential to characterize how attention behaves at extreme sequence lengths. W

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Towards Long-Horizon Vessel Trajectory and Destination Forecasting with Reasoning Large Language Models

DGX agent

arXiv:2606.08633v1 Announce Type: new Abstract: Long-horizon maritime trajectory prediction is important for shipping management, logistics planning, and maritime risk analysis, yet month-level foreca

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Towards Personalized Bangla Book Recommendation: A Large-Scale Heterogeneous Book Graph Dataset

DGX agent

arXiv:2602.12129v2 Announce Type: replace-cross Abstract: Personalized book recommendation in Bangla literature has been constrained by the lack of structured, large-scale, and publicly available data

model-releasesarxiv-cs-lg
9 Jun 2026
← Previous
1…145146147148149…361
Next →