AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,628 results
Model Releases

SA-Homo: Scale Adaptive Homography Estimation for Scale Variation Scenarios

DGX agent

arXiv:2606.30408v1 Announce Type: new Abstract: Homography estimation, as one of the fundamental problems in computer vision, remains challenged by scale variation scenarios where image pairs potentia

model-releasesarxiv-cs-cv
30 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

SABER-Math: Automated Benchmark for Information Retrieval Evaluation in Mathematics

DGX agent

arXiv:2606.29894v1 Announce Type: cross Abstract: As agentic AI systems tackle more complex mathematical tasks, they increasingly rely on information retrieval (IR) to search problem databases, theore

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

SADL: What to Ignore? A Benchmark for Subject-Aware Distractor Localization

DGX agent

arXiv:2606.30393v1 Announce Type: new Abstract: Photographs frequently contain visual distractors besides foregrounds and backgrounds of the intended subject, competing for attention and weakening com

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

SafePyramid: A Hierarchical Benchmark for In-context Policy Guardrailing

DGX agent

arXiv:2606.29887v1 Announce Type: new Abstract: In real-world applications, guardrails are often expected to identify unsafe user-model interactions according to application-specific safety policies,

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

SAKE: Software Architectural Knowledge Evaluation Benchmark for Large Language Models

DGX agent

arXiv:2606.29520v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used as assistants across the software development lifecycle, yet their ability to reason about software

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

SatSplat: Geometrically-Accurate Gaussian Splatting for Satellite Imagery

DGX agent

arXiv:2606.28581v1 Announce Type: new Abstract: High-resolution satellite imagery demands 3D reconstruction methods that deliver both speed and geometric accuracy. Recent adaptations of 3D Gaussian Sp

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

Scalar Representations of Neural Network Training Dynamics

DGX agent

arXiv:2606.30384v1 Announce Type: new Abstract: Training in artificial neural networks can be viewed as a trajectory evolving through a high-dimensional loss landscape. However, the large number of tr

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

ScAle: Attention Head Scaling as a Minimal Adapter for Spatial Reasoning in Vision Language Models

DGX agent

arXiv:2606.29579v1 Announce Type: cross Abstract: Spatial reasoning remains a persistent challenge for many vision language models (VLMs), and improving it typically requires fine-tuning with substant

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Scaling Textual Gradients via Sampling-Based Momentum

DGX agent

arXiv:2506.00400v4 Announce Type: replace-cross Abstract: LLM-based prompt optimization, which uses LLM-provided ``textual gradients'' (feedback) to refine prompts, has emerged as an effective method

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Scaling the Horizon, Not the Parameters: Reaching Trillion-Parameter Performance with a 35B Agent

DGX agent

arXiv:2606.30616v1 Announce Type: new Abstract: We introduce Agents-A1, a 35B Mixture-of-Experts Agentic Model that reaches trillion-parameter-level performance by scaling the agent horizon. We invest

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

SCARCE: Scalable Cascade Analysis for Rare-event Characterisation via Embeddings

DGX agent

arXiv:2606.29623v1 Announce Type: new Abstract: Rare events govern the safety profile of modern AI systems, yet their probabilities are extremely difficult to estimate: direct Monte Carlo requires pro

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

SciIR: A Large-scale Training Dataset and Benchmark for Scientific Image Reasoning Generation

DGX agent

arXiv:2606.30124v1 Announce Type: new Abstract: While Text-to-Image (T2I) models have shown remarkable success in generating photorealistic visual content, they still struggle with the rigorous semant

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

SciVisAgentBench: A Benchmark for Evaluating Scientific Data Analysis and Visualization Agents

DGX agent

arXiv:2603.29139v2 Announce Type: replace Abstract: Recent advances in large language models (LLMs) have enabled agentic systems to translate natural-language intent into executable scientific visuali

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Search for Truth from Reasoning: A Dynamic Representation Editing Framework for Steering LLM Trajectories

DGX agent

arXiv:2606.28589v1 Announce Type: new Abstract: Current approaches to enhance Large Language Model (LLM) reasoning, such as Chain-of-Thought and 'Wait' prompts, primarily encourage models to think mor

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

SEATauBench: Adapting Tool-Agent-User Evaluation Into Low-Resource Southeast Asian Languages

DGX agent

arXiv:2606.28715v1 Announce Type: cross Abstract: While AI development and evaluation for Southeast Asia (SEA) has grown rapidly, agent capabilities in regional languages are still poorly understood d

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Selective Memory Retention for Long-Horizon LLM Agents

DGX agent

arXiv:2606.29178v1 Announce Type: new Abstract: When does retention matter for memory-augmented LLM agents? We study this with TraceRetain, a lightweight framework for bounded external memory in froze

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Self-Supervised Theorem Discovery in a Formal Axiomatic System

DGX agent

arXiv:2606.28747v1 Announce Type: new Abstract: Recent artificial intelligence (AI) systems have shown remarkable progress in mathematical reasoning. Many existing approaches, including large language

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Semantic-Driven Scale and Spatial Selection for Efficient Cross-Modal Alignment in Referring Remote Sensing Image Segmentation

DGX agent

arXiv:2606.30244v1 Announce Type: new Abstract: Referring Remote Sensing Image Segmentation (RRSIS) seeks to localize and segment the target object or region specified by a natural language expression

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

SEO Agent. Security Agent. Design in Claude. Replit Desktop. Shopify on Replit. Skills + Custom Instructions. 450+ integrations. Package Fir…

DGX agent

SEO Agent. Security Agent. Design in Claude. Replit Desktop. Shopify on Replit. Skills + Custom Instructions. 450+ integrations. Package Firewall. And more! ✨ All shipped in June! 🤯 And there's a good

model-releasesreplit--x
30 Jun 2026
Model Releases

Sequential Hiring of Contingent Workers Through Learning-Based Optimization

DGX agent

arXiv:2606.18438v2 Announce Type: replace-cross Abstract: In this paper, we study a sequential workforce management problem in a contingent labor setting with uncertainty in both worker production and

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

Sequential Planning via Anchored Robotic Keypoints

DGX agent

arXiv:2606.30613v1 Announce Type: new Abstract: We present Sequential Planning via Anchored Robotic Keypoints, SPARK, a training-free neurosymbolic manipulation system that reaches 43.7% on six LIBERO

model-releasesarxiv-cs-ro
30 Jun 2026
Model Releases

SEVA: Self-Evolving Verification Agent with Process Reward for Fact Attribution

DGX agent

arXiv:2606.29713v1 Announce Type: cross Abstract: Hallucination is the reliability bottleneck for LLM-based agents, and fact attribution verifiers are the last line of defense -- yet today's verifiers

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

SFBench: The SciFy Scientific Feasibility Benchmark

DGX agent

arXiv:2606.29630v1 Announce Type: new Abstract: We present SFBench, a benchmark dataset for evaluating systems that assess the feasibility of scientific claims. SFBench includes 197 claims in material

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

SHOVIR: A Benchmark for Evaluating Vision Shortcut Learning in Radiology Report Generation

DGX agent

arXiv:2606.30201v1 Announce Type: cross Abstract: Current evaluation protocols for Vision-Language Models (VLMs) in Radiology Report Generation (RRG) rely on report-level metrics that measure lexical

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

Singular Learning and Occam's Razor in Deep Monomial Networks

DGX agent

arXiv:2606.28464v1 Announce Type: new Abstract: In the optimization of neural networks, gradient dynamics are influenced by critical points that arise from the model's architecture. These critical poi

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

SkelEM: Training-Signal Decoupling of Skeleton and Diffusion for Self-supervised Axial Super-Resolution in Volume Microscopy

DGX agent

arXiv:2606.30012v1 Announce Type: new Abstract: Volume microscopy, including electron and light microscopy, suffers from severe anisotropic resolution due to physical axial sectioning. Existing self-s

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

Sonnet 5 is here! This is going to support better long-running agents. Previous Sonnet models were unreliable, so it's great to see the impr…

DGX agent

Sonnet 5 is here! This is going to support better long-running agents. Previous Sonnet models were unreliable, so it's great to see the improved version that can complete agentic tasks more reliably.

model-releasesdair-ai--x
30 Jun 2026
Model Releases

Spatial Deconfounder: Interference-Aware Deconfounding for Spatial Causal Inference

DGX agent

arXiv:2510.08762v2 Announce Type: replace Abstract: Causal inference in spatial domains faces two intertwined challenges: (1) unmeasured spatial factors, such as weather, air pollution, or mobility, t

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

Spectral Embedding via Chebyshev Bases for Robust DeepONet Approximation

DGX agent

arXiv:2512.09165v2 Announce Type: replace Abstract: Deep Operator Networks (DeepONets) have emerged as a powerful framework for data-driven operator learning, providing flexible surrogates for nonline

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

SpreadsheetBench 2: Evaluating Agents on End-to-End Business Spreadsheet Workflows

DGX agent

arXiv:2606.29955v1 Announce Type: cross Abstract: Spreadsheets are widely used for business analysis, financial modeling, reporting, and decision-making. However, most existing spreadsheet benchmarks

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

SrDetection: A Self-Referential Framework for Data Leakage Detection in Code Large Language Models

DGX agent

arXiv:2606.29815v1 Announce Type: new Abstract: Evaluating code large language models (Code LLMs) requires reliable detection of data leakage, where benchmark performance is artificially inflated by e

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

Stability and Concentration in Nonlinear Inverse Problems with Block-Structured Parameters: Lipschitz Geometry, Identifiability, and an Application to Gaussian Splatting

DGX agent

arXiv:2602.09415v2 Announce Type: replace Abstract: We develop an operator-theoretic framework for stability and statistical concentration in nonlinear inverse problems with block-structured parameter

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

StarDojo: Benchmarking Open-Ended Behaviors of Agentic Multimodal LLMs in Production-Living Simulations with Stardew Valley

DGX agent

arXiv:2507.07445v3 Announce Type: replace Abstract: Autonomous agents navigating human society must master both production activities and social interactions, yet existing benchmarks rarely evaluate t

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Start building with Nano Banana 2 Lite and Gemini Omni Flash

DGX agent

Nano Banana 2 Lite is Google's fastest and most cost-efficient image generation model, generating images in approximately 4 seconds , while Gemini Omni Flash is a multimodal model supporting high-qual

model-releasesgoogle-deepmind
30 Jun 2026
Model Releases

Statistically Indistinguishable, Operationally Distinct: A Formal Barrier for Tabular Foundation Models

DGX agent

arXiv:2606.29091v1 Announce Type: cross Abstract: Tabular foundation models cannot reason about data produced by running systems without access to the rules that govern them. We make this statement fa

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Stay Unique, Stay Efficient: Preserving Model Personality in Multi-Task Merging

DGX agent

arXiv:2512.01461v2 Announce Type: replace-cross Abstract: Model merging has emerged as a promising paradigm for enabling multi-task capabilities without additional training. However, traditional basic

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

STEMGym: Benchmarking Sequential Decision-Making under Dose Budgets in Autonomous Electron Microscopy

DGX agent

arXiv:2606.29592v1 Announce Type: new Abstract: A central premise of autonomous scientific imaging is that smarter navigation, whether Bayesian, RL-based, or otherwise adaptive, is the principal lever

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

StrucTab: A Structured Optimization Framework for Table Parsing

DGX agent

arXiv:2606.29905v1 Announce Type: new Abstract: Table parsing aims to convert table images into structured, machine-readable representations, a task requiring the joint perception of complex spatial l

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

SurgVLA-Bench: Towards Evaluating Vision-Language-Action Models for Laparoscopic Surgical Robotics

DGX agent

arXiv:2606.29247v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models represent a promising direction for embodied intelligence in surgical robotics. Despite the prevalence of VLA benchm

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

SVC-Probe: A Framework for Evaluating Perturbation Generalization in Spatial Foundation-Model Embeddings

DGX agent

arXiv:2606.28465v1 Announce Type: cross Abstract: This work examines perturbation generalization in spatial foundation-model embeddings derived from fluorescence microscopy images. Although these mode

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

SVCBench: A Streaming Video Counting Benchmark for Spatial-Temporal State Maintenance

DGX agent

arXiv:2603.12703v3 Announce Type: replace Abstract: Video understanding requires models to continuously track and update world state during playback. Although existing benchmarks have advanced video u

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

SWE-fficiency: Can Language Models Optimize Real-World Repositories on Real Workloads?

DGX agent

arXiv:2511.06090v3 Announce Type: replace-cross Abstract: Optimizing the performance of large-scale software repositories demands expertise in code reasoning and software engineering (SWE) to reduce r

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

SWE-Together: Evaluating Coding Agents in Interactive User Sessions

DGX agent

arXiv:2606.29957v1 Announce Type: cross Abstract: Most coding-agent benchmarks are static: an agent receives a complete task description up front and is judged only by its final code. Real coding assi

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

SWITCH: Benchmarking Modeling and Handling of Tangible Interfaces in Long-horizon Embodied Scenarios

DGX agent

arXiv:2511.17649v4 Announce Type: replace-cross Abstract: Tangible control interfaces (TCIs), such as appliance panels, remotes, elevators, and embedded GUIs, are a fundamental component of everyday h

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Symbolic Mechanistic Data Attribution: Tracing Training Influence to Learned Behavioral Policies

DGX agent

arXiv:2606.29171v1 Announce Type: cross Abstract: While existing data attribution methods can identify which training examples build specific mechanistic circuits, they cannot explain how training dat

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

TextClusterLab: An Integrated Framework for Reliable Text Clustering Studies

DGX agent

arXiv:2606.28328v1 Announce Type: cross Abstract: In recent years, text clustering has become a critical technique for applications including intent discovery, topic mining, and recommendation systems

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

TF-MoE: Time-Frequency Mixture-of-Experts for Efficient Speech Separation

DGX agent

arXiv:2606.29575v1 Announce Type: cross Abstract: Recent advances in speech separation (SS) have led to compact front-end models with small parameter sizes, yet their high computational cost remains a

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Thank you to everyone to came to the Claude managed agents workshop at @aiDotEngineer with @gcemaj and I. We had an absolute blast sharing o…

DGX agent

Thank you to everyone to came to the Claude managed agents workshop at @aiDotEngineer with @gcemaj and I. We had an absolute blast sharing our journey and walking you through building your first agent

model-releasesswyx--x
30 Jun 2026
← Previous
1…154155156157158…472
Next →