AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Research

Semi-Supervised Sound Event Detection with Conditional Mixup and Embedding-Level Contrastive Loss

DGX agent

arXiv:2606.29901v1 Announce Type: cross Abstract: Sound event detection (SED) is a core module for acoustic environmental analysis, yet its performance is often limited by scarce labeled data. Recent

researcharxiv-cs-ai
30 Jun 2026
Agents
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

SemJoin: Semantic Join Optimization

DGX agent

arXiv:2606.29532v1 Announce Type: cross Abstract: Integrating unstructured data into relational database systems is increasingly important as demand grows for natural language querying and analysis. A

agentsarxiv-cs-ai
30 Jun 2026
Safety

Sequential Fairness Auditing with Limited Output Access

DGX agent

arXiv:2606.30338v1 Announce Type: new Abstract: External evaluations are becoming increasingly central to the governance of AI systems. In practice, however, independent auditors often have limited ac

safetyarxiv-cs-ai
30 Jun 2026
Research

Set-Inclusive Uncertainty Modeling for Robust Brain Tumor Segmentation

DGX agent

arXiv:2606.30374v1 Announce Type: cross Abstract: Multimodal MRI is essential for accurate brain tumor segmentation. However, acquiring all modalities at inference is often challenging in practice, wh

researcharxiv-cs-ai
30 Jun 2026
Model Releases

SEVA: Self-Evolving Verification Agent with Process Reward for Fact Attribution

DGX agent

arXiv:2606.29713v1 Announce Type: cross Abstract: Hallucination is the reliability bottleneck for LLM-based agents, and fact attribution verifiers are the last line of defense -- yet today's verifiers

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

SFBench: The SciFy Scientific Feasibility Benchmark

DGX agent

arXiv:2606.29630v1 Announce Type: new Abstract: We present SFBench, a benchmark dataset for evaluating systems that assess the feasibility of scientific claims. SFBench includes 197 claims in material

model-releasesarxiv-cs-ai
30 Jun 2026
Applications

SIMAX: A Scalable and Interpretable Framework for Multi-Fidelity and Annotated Clinician-Patient Dialogue Simulation

DGX agent

arXiv:2606.30491v1 Announce Type: cross Abstract: Background. The widespread deployment of ambient digital scribes is driving large-scale capture of clinician-patient dialogues. Human coding of clinic

applicationsarxiv-cs-ai
30 Jun 2026
Research

Situation Perception: A Necessary Primitive to Artificial Superintelligence

DGX agent

arXiv:2606.30481v1 Announce Type: cross Abstract: Current large language models are extraordinary statistical engines. They compress vast amounts of text into useful patterns and can explain science,

researcharxiv-cs-ai
30 Jun 2026
Research

Skin-R1: Clinical Knowledge-Guided Dermatological Diagnosis Using Vision-Language Models

DGX agent

arXiv:2511.14900v2 Announce Type: replace-cross Abstract: Vision--language models (VLMs) have recently shown promise for assisting clinical reasoning in dermatological diagnosis. However, their trustw

researcharxiv-cs-ai
30 Jun 2026
Applications

Solver-Verified Formulation Generation and Selection for Multi-Warehouse Inventory Allocation Using Large Language Models

DGX agent

arXiv:2606.29366v1 Announce Type: cross Abstract: Balance-oriented multi-warehouse inventory allocation is a recurring decision problem in large-scale e-commerce supply chains, in which a fixed replen

applicationsarxiv-cs-ai
30 Jun 2026
Local Ai

SonoCLIP: Mask-Guided Region-Aware Vision-Language Pretraining for Fetal Ultrasound Analysis

DGX agent

arXiv:2606.29586v1 Announce Type: cross Abstract: Vision-language foundation models have shown strong potential in medical image analysis. Although foundation models for ultrasound imaging have recent

local-aiarxiv-cs-ai
30 Jun 2026
Safety

SOTAlign: Semi-Supervised Alignment of Unimodal Vision and Language Models via Optimal Transport

DGX agent

arXiv:2602.23353v2 Announce Type: replace-cross Abstract: The Platonic Representation Hypothesis posits that neural networks trained on different modalities converge toward a shared statistical model

safetyarxiv-cs-ai
30 Jun 2026
Hardware

Spanning the Visual Analogy Space with a Weight Basis of LoRAs

DGX agent

arXiv:2602.15727v2 Announce Type: replace-cross Abstract: Visual analogy learning enables image editing via demonstration rather than textual description, allowing users to specify complex transformat

hardwarearxiv-cs-ai
30 Jun 2026
Local Ai

Spectral Perturbation of the Empirical Fisher Information Matrix under Weight Quantization

DGX agent

arXiv:2606.28432v1 Announce Type: cross Abstract: We study the spectral perturbation of the empirical Fisher Information Matrix (FIM) of a parametric statistical model under two structured perturbatio

local-aiarxiv-cs-ai
30 Jun 2026
Model Releases

SpreadsheetBench 2: Evaluating Agents on End-to-End Business Spreadsheet Workflows

DGX agent

arXiv:2606.29955v1 Announce Type: cross Abstract: Spreadsheets are widely used for business analysis, financial modeling, reporting, and decision-making. However, most existing spreadsheet benchmarks

model-releasesarxiv-cs-ai
30 Jun 2026
Hardware

SSM Meets Video Diffusion Models: Efficient Long-Term Video Generation with Structured State Spaces

DGX agent

arXiv:2403.07711v5 Announce Type: replace-cross Abstract: Given the remarkable achievements in image generation through diffusion models, the research community has shown increasing interest in extend

hardwarearxiv-cs-ai
30 Jun 2026
Research

Stabilizing Extrapolation in Looped Transformers via Learned Stochastic Stopping

DGX agent

arXiv:2606.29983v1 Announce Type: cross Abstract: Looped Transformers, which repeatedly apply a shared transformer block, are an architecturally natural fit for variable-length algorithmic tasks. Alth

researcharxiv-cs-ai
30 Jun 2026
Safety

StackingNet: Collective Inference Across Independent AI Foundation Models

DGX agent

arXiv:2602.13792v2 Announce Type: replace Abstract: Artificial intelligence built on large foundation models has transformed language understanding, computer vision, and reasoning, yet these systems r

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

StarDojo: Benchmarking Open-Ended Behaviors of Agentic Multimodal LLMs in Production-Living Simulations with Stardew Valley

DGX agent

arXiv:2507.07445v3 Announce Type: replace Abstract: Autonomous agents navigating human society must master both production activities and social interactions, yet existing benchmarks rarely evaluate t

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Statistically Indistinguishable, Operationally Distinct: A Formal Barrier for Tabular Foundation Models

DGX agent

arXiv:2606.29091v1 Announce Type: cross Abstract: Tabular foundation models cannot reason about data produced by running systems without access to the rules that govern them. We make this statement fa

model-releasesarxiv-cs-ai
30 Jun 2026
Research

Steerable Visual Representations

DGX agent

arXiv:2604.02327v2 Announce Type: replace-cross Abstract: Pretrained Vision Transformers (ViTs) such as DINOv2 and MAE provide generic image features that can be applied to a variety of downstream tas

researcharxiv-cs-ai
30 Jun 2026
Research

Structural Certification for Reliable Physical Design with Language Models

DGX agent

arXiv:2606.30107v1 Announce Type: new Abstract: An unreliable language model can be made to produce reliable physical designs if the authority to assert is moved out of the model: the model proposes,

researcharxiv-cs-ai
30 Jun 2026
Applications

SUMO: Segment and Track Any Motion with Nonlinear State Space Models

DGX agent

arXiv:2606.29861v1 Announce Type: cross Abstract: Visual Object Tracking (VOT) and Moving Object Segmentation (MOS) are two fundamental tasks in computer vision that involve both spatial and temporal

applicationsarxiv-cs-ai
30 Jun 2026
Model Releases

SurgVLA-Bench: Towards Evaluating Vision-Language-Action Models for Laparoscopic Surgical Robotics

DGX agent

arXiv:2606.29247v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models represent a promising direction for embodied intelligence in surgical robotics. Despite the prevalence of VLA benchm

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

SVC-Probe: A Framework for Evaluating Perturbation Generalization in Spatial Foundation-Model Embeddings

DGX agent

arXiv:2606.28465v1 Announce Type: cross Abstract: This work examines perturbation generalization in spatial foundation-model embeddings derived from fluorescence microscopy images. Although these mode

model-releasesarxiv-cs-ai
30 Jun 2026
Hardware

SwarmX: Agentic Scheduling for Low-Latency Agentic Systems

DGX agent

arXiv:2606.21401v2 Announce Type: replace-cross Abstract: Agentic AI applications compose multiple model calls and tool executions, creating new scheduling challenges for GPU-CPU clusters. Their infer

hardwarearxiv-cs-ai
30 Jun 2026
Model Releases

SWE-fficiency: Can Language Models Optimize Real-World Repositories on Real Workloads?

DGX agent

arXiv:2511.06090v3 Announce Type: replace-cross Abstract: Optimizing the performance of large-scale software repositories demands expertise in code reasoning and software engineering (SWE) to reduce r

model-releasesarxiv-cs-ai
30 Jun 2026
Tutorials

SWE-MeM: Learning Adaptive Memory Management for Long-Horizon Coding Agents

DGX agent

arXiv:2606.28434v1 Announce Type: cross Abstract: Long-horizon software engineering agents often need to manage lengthy and noisy interaction histories under limited context budgets. Existing memory m

tutorialsarxiv-cs-ai
30 Jun 2026
Model Releases

SWE-Together: Evaluating Coding Agents in Interactive User Sessions

DGX agent

arXiv:2606.29957v1 Announce Type: cross Abstract: Most coding-agent benchmarks are static: an agent receives a complete task description up front and is judged only by its final code. Real coding assi

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

SWITCH: Benchmarking Modeling and Handling of Tangible Interfaces in Long-horizon Embodied Scenarios

DGX agent

arXiv:2511.17649v4 Announce Type: replace-cross Abstract: Tangible control interfaces (TCIs), such as appliance panels, remotes, elevators, and embedded GUIs, are a fundamental component of everyday h

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Symbolic Mechanistic Data Attribution: Tracing Training Influence to Learned Behavioral Policies

DGX agent

arXiv:2606.29171v1 Announce Type: cross Abstract: While existing data attribution methods can identify which training examples build specific mechanistic circuits, they cannot explain how training dat

model-releasesarxiv-cs-ai
30 Jun 2026
Applications

T3R: Deeper Test-Time Adaptation for Graph Neural Networks via Gradient Rotation

DGX agent

arXiv:2606.30011v1 Announce Type: cross Abstract: Graph Neural Networks (GNNs) deployed in real-world systems typically have fixed weights, often leading to degraded performance under distribution shi

applicationsarxiv-cs-ai
30 Jun 2026
Research

Tactile Gesture Recognition with Built-in Joint Sensors for Industrial Robots

DGX agent

arXiv:2508.12435v2 Announce Type: replace-cross Abstract: While gesture recognition using vision or robot skins is an active research area in Human-Robot Collaboration (HRC), this paper explores deep

researcharxiv-cs-ai
30 Jun 2026
Local Ai

TAR: Temporal Anchor-Constrained Reasoning for Video Temporal Grounding

DGX agent

arXiv:2508.07683v2 Announce Type: replace-cross Abstract: Video Temporal Grounding (VTG) aims to localize specific video segments corresponding to natural language queries. While recent Large Vision-L

local-aiarxiv-cs-ai
30 Jun 2026
Tutorials

Temporal Feature Extractors in EEG Foundation Models: A Controlled Comparison Including a Pretrained Time-Series Model

DGX agent

arXiv:2606.30104v1 Announce Type: new Abstract: Electroencephalography (EEG) foundation models aim to learn generalizable representations from large-scale brain recordings. However, the role of tempor

tutorialsarxiv-cs-ai
30 Jun 2026
Safety

TERC: A Transfer Entropy Redundancy Criterion for State Variable Selection in Reinforcement Learning

DGX agent

arXiv:2401.11512v2 Announce Type: replace-cross Abstract: Identifying the most suitable variables to represent the state is a fundamental challenge in Reinforcement Learning (RL). These variables must

safetyarxiv-cs-ai
30 Jun 2026
Safety

Test-Time Detoxification without Training or Learning Anything

DGX agent

arXiv:2602.02498v2 Announce Type: replace-cross Abstract: Large language models can produce toxic or inappropriate text even for benign inputs, creating risks when deployed at scale. Detoxification is

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

TF-MoE: Time-Frequency Mixture-of-Experts for Efficient Speech Separation

DGX agent

arXiv:2606.29575v1 Announce Type: cross Abstract: Recent advances in speech separation (SS) have led to compact front-end models with small parameter sizes, yet their high computational cost remains a

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

The Complexity Ceiling Benchmark: A Multi-Domain Evaluation of Sequential Reasoning Under Depth Scaling

DGX agent

arXiv:2606.29278v1 Announce Type: new Abstract: We introduce the Complexity Ceiling Benchmark (CCB), a controlled evaluation of how language-model reasoning decays as the number of required sequential

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

The CRISTAL Method: Neurosymbolic analysis from AI-synthesized world models

DGX agent

arXiv:2606.29799v1 Announce Type: new Abstract: This project introduces the CRISTAL Method (Coherent Reliable Intentional Synthesis of Truthful Analysis Logic), a neurosymbolic framework for automatin

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

The Crowded Embedding Space: A Mean-Field Mechanism for Emergent Marginalization in Retrieval-Augmented Agents

DGX agent

arXiv:2606.28343v1 Announce Type: cross Abstract: Retrieval-augmented generative agents rely on retrieval for grounding, yet are typically evaluated on a query-by-query basis. This isolates interactio

safetyarxiv-cs-ai
30 Jun 2026
Agents

The Emergence of Autonomous Penetration Capabilities in Large Language Model-Powered AI Systems

DGX agent

arXiv:2606.13079v2 Announce Type: replace-cross Abstract: Nowadays, the autonomous execution of cyberattacks capable of causing substantial real-world harm is widely regarded as one of the critical re

agentsarxiv-cs-ai
30 Jun 2026
Model Releases

The FIL Hypothesis: Inductive Biases Help with Kernel Engineering

DGX agent

arXiv:2606.30442v1 Announce Type: new Abstract: The Bitter Lesson, which posits that general-purpose methods that scale with computation and data ultimately outperform those with built-in human knowle

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

The Heterogeneous Safety Impacts of Benign Multilingual Fine-Tuning

DGX agent

arXiv:2606.28843v1 Announce Type: cross Abstract: Fine-tuning a large language model is a ubiquitous method for enhancing its capability on a specific downstream task. However, prior work has shown th

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

The Hidden Cost of Structured Generation in LLMs: Draft-Conditioned Constrained Decoding

DGX agent

arXiv:2603.03305v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used to generate executable outputs, JSON objects, and API calls, where a single syntax error ca

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

The Human Creativity Benchmark

DGX agent

arXiv:2606.30561v1 Announce Type: new Abstract: Modern AI evaluation frameworks treat evaluator disagreement as noise to be resolved. In creative domains, professional disagreement reflects genuine di

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

The Interference Gap: Comparing Retrieval Bounds in Human Memory and RAG Systems

DGX agent

arXiv:2606.28327v1 Announce Type: cross Abstract: How do retrieval bounds compare between human episodic memory and Retrieval-Augmented Generation (RAG) systems under semantic interference? We present

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

The Joint Effect of Quantization and Sampling Temperature on LLM Safety Alignment: A Factorial Analysis

DGX agent

arXiv:2606.29581v1 Announce Type: cross Abstract: Modern LLM deployments routinely compress models and raise sampling temperature to reduce cost, latency, or repetition, yet safety evaluations usually

safetyarxiv-cs-ai
30 Jun 2026
← Previous
1…144145146147148…448
Next →