AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
86,428Total entries
1Added by human
86,427Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
23,185 results
Model Releases

Remember with Confidence: Uncertainty Quantification for Spatio-temporal Memory with Probabilistic Guarantees

DGX agent

arXiv:2606.08277v1 Announce Type: new Abstract: Long-horizon robot operation requires spatio-temporal memory to record the environment state and recall it for downstream reasoning. Scene graphs and re

model-releasesarxiv-cs-cv
9 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Repetition Mismatch: Why Data Mixture Experiments Don't Scale and How to Fix Them

DGX agent

arXiv:2606.07597v1 Announce Type: cross Abstract: Pre-training data mixtures are commonly tuned by running small-scale experiments and extrapolating to the target training budget. When high-quality da

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Report: GKE Inference Gateway delivers up to 92% faster AI responses

DGX agent

As generative AI moves from experimental pilots to massive production environments, the efficiency of your infrastructure becomes the ultimate differentiator. One way to get the most out of it and min

model-releasesgoogle-cloud-ai
9 Jun 2026
Model Releases

ResearchClawBench: A Benchmark for End-to-End Autonomous Scientific Research

DGX agent

arXiv:2606.07591v1 Announce Type: cross Abstract: AI coding agents are increasingly used for scientific work, but their end-to-end autonomous research capability remains difficult to verify. We presen

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Revisiting Training Scale: An Empirical Study of Token Count, Power Consumption, and Parameter Efficiency

DGX agent

arXiv:2601.06649v2 Announce Type: replace-cross Abstract: Research in machine learning has questioned whether increases in training token counts reliably produce proportional performance gains in larg

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

RiskNet: A large-scale dataset of AI risk incidents from news with alignment and multi-dimensional annotations

DGX agent

arXiv:2606.08376v1 Announce Type: cross Abstract: As artificial intelligence (AI) systems are increasingly deployed across socially consequential domains, reports of AI-related harms and failures have

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Robust-U1: Can MLLMs Self-Recover Corrupted Visual Content for Robust Understanding?

DGX agent

arXiv:2606.08063v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have demonstrated remarkable success in visual understanding, yet their performance degrades significantly un

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Rosetta Memory: Adaptive Memory for Cross-LLM Agents

DGX agent

arXiv:2606.07711v1 Announce Type: cross Abstract: Memory is the key component for transforming a stateless LLM into a persistent, evolving agent through experience accumulation, long-horizon planning,

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

RTL-BenchLS: A Large-Scale Benchmark for RTL Reasoning and Generation with Large Language Models

DGX agent

arXiv:2606.08976v1 Announce Type: new Abstract: LLM-based RTL generation and reasoning is a promising direction for hardware design automation. High-quality benchmarks are critical infrastructure for

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Rubrik turns its platform into an AI agent and ships Agent Cloud for Claude

DGX agent

Rubrik Inc. today turned its data security platform into an autonomous agent and made its control layer for Anthropic PBC’s Claude generally available, the headline items in a wave of announcements at

model-releasessiliconangle
9 Jun 2026
Model Releases

RunAgent SuperBrowser: A Theory of Autonomous Web Navigation Grounded in Human Browsing Behaviour

DGX agent

arXiv:2606.09399v1 Announce Type: new Abstract: We present SUPERBROWSER, an autonomous web-navigation agent designed against a single guiding hypothesis: a web agent should browse the way a person bro

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Safe-RULE: Safe Reinforcement UnLEarning

DGX agent

arXiv:2606.09559v1 Announce Type: cross Abstract: Offline safe reinforcement learning (Safe RL) enables policy learning without online interactions, making it suitable for safety-critical systems such

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

SafeRun: Enabling Determinism in LLM Planning for Running

DGX agent

arXiv:2606.09027v1 Announce Type: cross Abstract: Large Language Models enable flexible natural-language planning but remain unreliable in determinism-critical domains due to their probabilistic natur

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

SC3: The Multi-Solvent Solubility Challenge and Benchmark

DGX agent

arXiv:2606.07656v1 Announce Type: cross Abstract: Solubility prediction is a standard benchmark in computational chemistry, yet multi-solvent models which reportedly approach the experimental-noise ce

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Scaffold Effects on GAIA: A Controlled Comparison

DGX agent

arXiv:2606.08529v1 Announce Type: new Abstract: Published agent capability scores conflate what a model can do with what its scaffold lets it do, and the magnitude of this elicitation gap is not well

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

ScaleSweep: Accurate NVFP4 Post-Training Quantization of LLMs via Block Scale Initialization

DGX agent

arXiv:2606.07618v1 Announce Type: cross Abstract: NVFP4 is a recently introduced hardware-supported FP4 format that improves the fidelity of 4-bit quantization through fine-grained block scales. Howev

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Scaling by Diversified Experience for Vision-Language-Action Models

DGX agent

arXiv:2606.09009v1 Announce Type: new Abstract: Vision-Language-Action models face significant challenges in real-world deployment due to the entanglement of high-level reasoning with low-level contro

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Scaling Laws for Masked-Reconstruction Transformers on Single-Cell Transcriptomics

DGX agent

arXiv:2602.15253v2 Announce Type: replace Abstract: Neural scaling laws -- power-law relationships between loss, model size, and data -- have been extensively documented for language and vision transf

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

scCBGM: Interpretable Single-Cell Counterfactual Editing

DGX agent

arXiv:2606.07760v1 Announce Type: new Abstract: Understanding cellular phenotypes and how they respond to perturbations is critical for disease biology and therapeutic design. Single-cell RNA sequenci

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

SceneConductor: 3D Scene Generation from Single Image with Multi-Agent Orchestration

DGX agent

arXiv:2606.08402v1 Announce Type: cross Abstract: Generating complete 3D scenes from a single image requires inferring globally consistent geometry, object relationships, and environmental context fro

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Sci-Rho: A Multilingual Visually-Grounded Symbolic Benchmark for STEM Problems

DGX agent

arXiv:2606.08034v1 Announce Type: cross Abstract: Symbolic benchmarks have emerged as a key approach to assess model robustness under minor modifications to STEM-related questions. However, existing s

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

SciFlow-Bench: Evaluating Structure-Aware Scientific Diagram Generation via Inverse Parsing

DGX agent

arXiv:2602.09809v2 Announce Type: replace Abstract: Scientific diagrams convey explicit structural information, yet modern text-to-image models often produce visually plausible but structurally incorr

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

See how Claude Fable 5 compares across every model: http://cursor.com/evals

DGX agent

Claude Fable 5 is compared against other AI models on various evaluation metrics through Cursor's benchmarking tool. The evaluation likely covers performance across different tasks such as coding, rea

model-releasescursor--x
9 Jun 2026
Model Releases

See More, Think Deeper: Query-Expanded Visual Evidence and Answer-Clue Guided Reflection for Long Video Understanding

DGX agent

arXiv:2606.09064v1 Announce Type: cross Abstract: Recent advances in Video Large Language Models (Video-LLMs) have enabled performance on long-video understanding tasks. However, existing methods stil

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

SegmentAnyTreeV2: Scaling Transformer-Based Tree Instance Segmentation Across Sensors, Platforms, and Forests

DGX agent

arXiv:2606.08206v1 Announce Type: new Abstract: We present SegmentAnyTreeV2, a sensor- and platform-agnostic framework for semantic and instance segmentation of forest point clouds. The model combines

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Semi-supervised Source Detection in Astronomical Images: New Benchmark and Strong Baseline

DGX agent

arXiv:2606.09219v1 Announce Type: new Abstract: Source detection in modern observational astronomy is a cornerstone for localizing and identifying stellar sources accurately. It is crucial for studies

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

SENTRY: Statistical Reliability Analysis of Vision Transformers Under Soft Errors

DGX agent

arXiv:2606.07620v1 Announce Type: cross Abstract: With the growth of Vision Transformers in safety-critical domains like autonomous systems and medical imaging, ensuring their reliability against soft

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Seq103: A Unified Neuroevolution Framework for Compact Sequence Architecture Discovery

DGX agent

arXiv:2606.07664v1 Announce Type: cross Abstract: Neuroevolution is a representative neural architecture search paradigm that evolves both network topology and weights through evolutionary algorithms.

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Setting a custom price for a model in AgentsView

DGX agent

TIL: Setting a custom price for a model in AgentsView I've been really enjoying AgentsView by Wes McKinney as a tool for exploring my token usage across different coding agents running on my laptop. C

model-releasessimon-willison
9 Jun 2026
Model Releases

Shared Latent Structures Enable Unified Backdoor Detection and Mitigation in LLMs

DGX agent

arXiv:2606.07963v1 Announce Type: new Abstract: Backdoor attacks in large language models (LLMs) are often treated as isolated trigger-response failures, motivating defenses tailored to specific trigg

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Shift-Dependent Asymmetry: Orthogonal Inverse Low-Rank Adaptation for Federated Medical Segmentation

DGX agent

arXiv:2606.08687v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) enables efficient federated fine-tuning of segmentation foundation models for medical imaging. However, most federated LoRA m

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Signals Are Not States: Neuro-Symbolic Safeguards for Culturally Aware Classroom AI

DGX agent

arXiv:2603.22793v2 Announce Type: replace Abstract: Classroom AI systems increasingly infer high-level educational states such as engagement, confusion, collaboration, participation, and instructional

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

SIMPLE: Simulation-Based Policy Learning and Evaluation for Humanoid Loco-manipulation

DGX agent

arXiv:2606.08278v1 Announce Type: new Abstract: Humanoid foundation models are advancing faster than we can evaluate them. While real-world testing is expensive and difficult to reproduce, existing si

model-releasesarxiv-cs-ro
9 Jun 2026
Model Releases

SLMJury: Can Small Language Models Judge as Well as Large Ones?

DGX agent

arXiv:2606.07810v1 Announce Type: cross Abstract: Large language models (LLMs) are widely used as judges for evaluating model outputs, but their high cost, latency, and opacity limit scalability. We i

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

SNN-MLIR: An MLIR Dialect for Compiling Neuromorphic SNNs from NIR to Bare-Metal C

DGX agent

arXiv:2606.09213v1 Announce Type: cross Abstract: Spiking neural networks (SNNs) are increasingly trained in a wide range of frameworks (SnnTorch, Lava, Norse, and others) each with its own model form

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

SoK: Reconstruction Attacks on Synthetic Tabular Data (Insights from Winning the NIST CRC)

DGX agent

arXiv:2606.08372v1 Announce Type: cross Abstract: Synthetic data is increasingly promoted as a privacy-preserving substitute for releasing sensitive tabular records, yet its central adversarial threat

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Solving Inverse Problems with Flow-based Models via Model Predictive Control

DGX agent

arXiv:2601.23231v2 Announce Type: replace-cross Abstract: Flow-based generative models provide strong unconditional priors for inverse problems, but guiding their dynamics for conditional generation r

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Sovereign AI for all.

DGX agent

Cohere advocates for democratizing access to sovereign AI systems, enabling organizations and nations to develop and deploy their own AI models independently rather than relying on centralized provide

model-releasescohere--x
9 Jun 2026
Model Releases

Sparse Autoencoders Reveal Interpretable and Steerable Features in VLA Models

DGX agent

arXiv:2603.19183v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models have emerged as a promising approach for general-purpose robot manipulation. However, little research has mechan

model-releasesarxiv-cs-ro
9 Jun 2026
Model Releases

SpatialWorld: Benchmarking Interactive Spatial Reasoning of Multimodal Agents in Real-World Tasks

DGX agent

arXiv:2606.09669v1 Announce Type: new Abstract: Spatial reasoning is a foundational capability for multimodal large language models (MLLMs) to perceive and operate within the physical world. However,

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

SpectrumKV: Per-Token Mixed-Precision KV Cache Transfer for Prefill-Decode Disaggregated LLM Serving

DGX agent

arXiv:2606.08635v1 Announce Type: new Abstract: Prefill-decode (PD) disaggregation decouples prompt processing from token generation, but it also turns the key-value (KV) cache into a network payload.

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

SSAFE: Simple and Strong AI-Generated Image Detection via Frozen Vision Encoders

DGX agent

arXiv:2606.08634v1 Announce Type: new Abstract: The rapid advancement of generative models has blurred the boundary between synthetic and real imagery, creating an urgent need for reliable deepfake de

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Stabilizing On-Policy Distillation for MLLM Reasoning with Global Normalization

DGX agent

arXiv:2606.09091v1 Announce Type: cross Abstract: On-policy distillation (OPD) has recently emerged as an important post-training paradigm. By using a stronger teacher model to provide dense, fine-gra

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Still: Amortized KV Cache Compaction in a Single Forward Pass

DGX agent

arXiv:2606.07878v1 Announce Type: new Abstract: The KV cache is the memory bottleneck of long-horizon language model deployment. Practically, a deployable compactor must be lightweight enough to call

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Storage Insights datasets: Enabling org-wide operational discovery with activity insights

DGX agent

As enterprise storage footprints scale to billions of objects, AI applications and agentic workloads are fundamentally shifting the role of storage from a passive repository to the foundation of the d

model-releasesgoogle-cloud-ai
9 Jun 2026
Model Releases

Strained Coherence: A Pre-Failure Signal in Coding Agent Execution Trajectories

DGX agent

arXiv:2606.07889v1 Announce Type: cross Abstract: LLM-based coding agents sometimes acknowledge a problem in their own reasoning and then proceed anyway. We call this pattern strained coherence: a saf

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Streaming Interventions: Can Video Large Language Models Correct Mistakes as They Occur?

DGX agent

arXiv:2606.09547v1 Announce Type: new Abstract: Learning everyday skills, like cooking a dish, relies increasingly on instructional media such as online videos. This opens the door to the use of video

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Stress-testing medical large language models reveals latent safety pathology beyond benchmark accuracy

DGX agent

arXiv:2606.07929v1 Announce Type: new Abstract: Large language models (LLMs) are entering clinical practice based on benchmark accuracy that may fail to detect safety-relevant failure modes. Here we p

model-releasesarxiv-cs-ai
9 Jun 2026
← Previous
1…211212213214215…484
Next →