AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
Human
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
84,647 results
28 Jul 2026

Sources: the OpenAI agent that breached Hugging Face also compromised a customer at AI infrastructure company Modal Labs (Reuters)

AgentsDGX agent

Reuters: Sources: the OpenAI agent that breached Hugging Face also compromised a customer at AI infrastructure company Modal Labs — The rogue agent that escaped from OpenAI and went on a days-long hac

Sparse Autoencoders Encode Both Concepts and Functions: The Downstream Geometry of Feature Effects

ResearchDGX agent

arXiv:2607.24645v1 Announce Type: cross Abstract: The wide-scale use of sparse autoencoders (SAEs) as interpretability tools is limited by inconsistent links between SAE features and model behavior. F

Sparse Evidence Can Suffice: Agentic Evidence Seeking for Multimodal Video Misinformation Detection

AgentsDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.18080v2 Announce Type: replace-cross Abstract: Multimodal video misinformation detection is commonly formulated as a holistic video-understanding task, where the entire video and its associ

Sparse Gaussian-Mixture-Model Q-Functions via Hadamard Overparametrization for Online Reinforcement Learning

Model ReleasesDGX agent

arXiv:2607.23474v1 Announce Type: new Abstract: This paper develops an online, off-policy policy-iteration framework for reinforcement learning (RL), based on sparse Gaussian-mixture-model Q-functions

Spatial-IQ: Deconstructing Spatial Intelligence via Hierarchical Capability Tests

HardwareDGX agent

arXiv:2607.22864v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) excel at visual interpretation but fail on spatial reasoning tasks that humans solve reliably. Existing bench

Spatial Prediction of Soil Microplastics and Organic Matter Using Graph Attention Networks

Local AiDGX agent

arXiv:2607.22875v1 Announce Type: cross Abstract: Accurate estimation of soil microplastics and organic matter is essential to assess ecosystem health and support sustainable land use. This study pres

Spatial Reasoning in LLM Game Agents: Impact of Causal Context and Multi-Step Planning

Model ReleasesDGX agent

arXiv:2607.22732v1 Announce Type: new Abstract: LLM-based game agents often perform poorly on more complex tasks. This work examines whether these failures are linked to limited spatial reasoning and

Spatially-Aware Class-Agnostic Object Counting

ResearchDGX agent

arXiv:2607.16826v2 Announce Type: replace Abstract: Generalised object counting aims to estimate the number of instances of an arbitrary object category from a single image, but many recent methods ca

Spatio-Temporal Conditional Denoising Transformer for Modality-Missing RGBT Tracking

Model ReleasesDGX agent

arXiv:2607.24701v1 Announce Type: new Abstract: Missing modalities in RGBT tracking often lead to incomplete and unstable multimodal feature representations that greatly degrade the performance. Exist

Spatula: Exploring On-Demand In-Situ Interfaces and Interaction for Attribute Control

Model ReleasesDGX agent

arXiv:2607.10405v2 Announce Type: replace-cross Abstract: Controlling attributes is a critical step toward achieving the final creative outcome, yet current approaches fall short in supporting users i

spec: add DSpark speculative decoding by wjinxu · Pull Request #25173 · ggml-org/llama.cpp

Model ReleasesDGX agent

It's time to experiment using DSpark! Please share your stats(pp/tg improvements). DSpark related stuff to check: DeepSpec - a deepseek-ai Collection DeepSeek-V4 with DSpark - DeepSeek-V4-Pro-DSpark &

SpecAHD: Localize to Specialize for Automated Heuristic Design in Large-Scale Routing Problems

Local AiDGX agent

arXiv:2607.23676v1 Announce Type: new Abstract: LLM-based automated heuristic design (AHD) typically scores executable programs on complete instances or within fixed solver components. In large-scale

SpecBox: Speculative Sandbox Scheduling for Efficient LLM Agent Serving

AgentsDGX agent

arXiv:2607.23933v1 Announce Type: cross Abstract: As LLM agents increasingly rely on the Model Context Protocol (MCP) to invoke isolated external sandboxes, disaggregated sandbox deployment introduces

SpecFormer: Mitigating Embedding and Attention Collapse via Spectral-Aware Transformer for Recommendation

SafetyDGX agent

arXiv:2607.24025v1 Announce Type: cross Abstract: Transformer architectures have achieved remarkable success across diverse domains; however, directly applying their standard self-attention mechanism

Spectral-Aware Analytic Class-Incremental Learning for Long-Tailed Distributions

ResearchDGX agent

arXiv:2607.22931v1 Announce Type: cross Abstract: Analytic Continual Learning (ACL) offers a computationally efficient alternative to gradient-based approaches. Recent ACL methods are based on Recursi

Spectral Dynamics of Semantic Drift in Clinical Multi-Agent Language Model Networks

SafetyDGX agent

arXiv:2607.22758v1 Announce Type: cross Abstract: The integration of iterative LLMs within multi-agent diagnostic frameworks requires a rigorous quantitative reevaluation of underlying communication t

Speech Signals Complement LLMs for Predicting Interpersonal Attraction in Speed Dating

ResearchDGX agent

arXiv:2607.23037v1 Announce Type: new Abstract: Large language models (LLMs) can predict interpersonal attraction from conversation transcripts, but it remains unclear what a speech predictor can add

Speed Reading Tool Powered by Artificial Intelligence for Students with ADHD, Dyslexia, and Short Attention Span

Model ReleasesDGX agent

arXiv:2307.14544v2 Announce Type: replace-cross Abstract: This paper presents an artificial intelligence tool designed to assist students with dyslexia, ADHD, and short attention spans in processing t

SPRKD: Effective Knowledge Distillation for Deep Neural Networks via Saddle Region Approximation

Model ReleasesDGX agent

arXiv:2607.23346v1 Announce Type: new Abstract: Modern deep neural networks are potent catalysts for scientific and industrial impact, yet excessive parameter counts impede deployment in low-compute s

SQBench: A Benchmark for Evaluating Task Delivery by Language-Model Agents in Production-Oriented Workflows

Model ReleasesDGX agent

arXiv:2607.23123v1 Announce Type: new Abstract: Existing evaluations of large language models cover knowledge, reasoning, coding, and tool use, but they rarely treat a verifiable deliverable produced

Stability of AI Governance Systems: A Coupled Dynamics Model of Public Trust and Social Disruptions

Model ReleasesDGX agent

arXiv:2603.20248v2 Announce Type: replace-cross Abstract: AI systems are increasingly entrenched in public governance, yet scholarship lacks formal tools to determine when deviations of public trust i

Stabilizing Deep Reconstruction Operators with Contractive Anchoring

ResearchDGX agent

arXiv:2607.23341v1 Announce Type: cross Abstract: Pretrained deep denoisers can be used to solve a wide range of model-based image reconstruction tasks via Plug-and-Play (PnP) and Regularization-by-De

Stacking the Deck: Tunable Trainability in Stacked LCUs

ResearchDGX agent

arXiv:2607.24686v1 Announce Type: cross Abstract: Variational quantum circuits have been central to many proposed near-term applications of quantum computing, but a growing body of evidence suggests t

StageGuard: Physiologically Constrained Sleep Staging

SafetyDGX agent

arXiv:2607.23284v1 Announce Type: new Abstract: Automated sleep staging is increasingly used in large-scale studies to derive sleep-architecture endpoints: total sleep time, REM latency, sleep efficie

STAIF: A Stage-wise Optimization for Complex Instruction Following

SafetyDGX agent

arXiv:2607.22649v1 Announce Type: new Abstract: Following complex instructions with multiple explicit constraints remains a fundamental challenge for large language models (LLMs). Existing alignment m

StanceBench: A Benchmark for Audio LLM-Based Interpersonal Stance Evaluation from Speech

Model ReleasesDGX agent

arXiv:2607.22658v1 Announce Type: new Abstract: Speech-to-speech dialogue models increasingly depend on prosody and interactional nuance to convey social intent, yet benchmarks for these cues remain l

StanceFlip: A Comprehensive Multi-Dimensional Benchmark for Multimodal Conversational Stance Flipping Forecasting

Model ReleasesDGX agent

arXiv:2607.24191v1 Announce Type: cross Abstract: Conversational stance detection has shifted from static text analysis to dynamic multimodal modeling. However, existing benchmarks exhibit three key l

StAR: Segment Anything Reasoner

Model ReleasesDGX agent

arXiv:2603.14382v2 Announce Type: replace Abstract: As AI systems are being integrated more rapidly into diverse and complex real-world environments, the ability to perform holistic reasoning over an

StateAct: Program State, before Pixels, for Long-Horizon Computer-Use Agents

Model ReleasesDGX agent

arXiv:2607.22798v1 Announce Type: cross Abstract: Computer-use agents are usually improved by strengthening perception: better models for reading a screenshot and choosing where to click. Yet a screen

Statistically Supported LLM Ingredient and Recipe Data Collection in Computational Nutrition

ResearchDGX agent

arXiv:2607.23273v1 Announce Type: cross Abstract: Computational nutrition needs precise ingredient data, but current databases are incomplete, inconsistent, and built for human reference rather than a

STEER: Steerable Dyadic Head Avatars

ResearchDGX agent

arXiv:2607.23840v1 Announce Type: new Abstract: Facial movement and expression are central to face-to-face communication, conveying turn-taking, attention, agreement, and engagement alongside speech.

Steerable Chatbots: Exploring Personalization Control Interfaces via LLM Activation Steering

ResearchDGX agent

arXiv:2505.04260v3 Announce Type: replace-cross Abstract: Personalizing LLM responses typically requires users to articulate their preferences through prompting, which can be burdensome at cold start

StepX-Edge: An On-Device UI Vision-Language Model via Architecture-Training-Deployment Co-Design

Model ReleasesDGX agent

arXiv:2607.22708v1 Announce Type: new Abstract: Deploying a vision-language model with full UI understanding on end devices has long been trapped between accuracy and efficiency: on one side is the ac

Stochastic Counterdiabatic Driving via Biorthogonal Liouvillian Eigenmodes

ResearchDGX agent

arXiv:2607.24393v1 Announce Type: cross Abstract: Finite-time driving of stochastic systems generates excess dissipation, causing the evolving probability distribution to lag behind the instantaneous

Stress-Testing EEG Foundation Models for Clinical Decoding: Dataset Identity and Targeted Negative Controls

Model ReleasesDGX agent

arXiv:2607.24519v1 Announce Type: cross Abstract: Pretrained EEG foundation models are increasingly proposed for clinical decoding, but their transfer across populations and robustness to negative con

Stress-testing large language model agents in a robotic chemistry laboratory

AgentsDGX agent

arXiv:2607.23045v1 Announce Type: new Abstract: AI is evaluated through knowledge, reasoning and plan generation, yet scientific agency requires reliable physical action and adaptation to evidence. He

Structural Loss Metrics for Tensor Approximation via Matrix Low-Rank Approximation

ResearchDGX agent

arXiv:2607.24009v1 Announce Type: new Abstract: Matricized low-rank approximation via SVD is a standard surrogate for tensor decompositions, but entry-wise reconstruction error fails to capture multiw

Structural Preservation Governs Data Augmentation in Deep Learning-Based Laser Speckle Material Classification

ResearchDGX agent

arXiv:2607.22725v1 Announce Type: cross Abstract: Data augmentation is routinely used to improve generalization in image classification, but the assumptions underlying standard policies are poorly mat

Structure over Depth: A Single-Block Spatio-Temporal Transformer for Multi-Entity Reasoning

TutorialsDGX agent

arXiv:2607.23077v1 Announce Type: new Abstract: Modeling multi-entity temporal data requires capturing dependencies across entities, time, and their interactions. Transformer-based approaches perform

Structure Over Scale: Schema-Constrained Causal Graphs for RAG

ResearchDGX agent

arXiv:2607.22592v1 Announce Type: new Abstract: Graph-based retrieval-augmented generation (GraphRAG) grounds answers in structured knowledge, but current systems extract entities and relationships ex

Structured Observation Language for Efficient and Generalizable Vision-Language Navigation

AgentsDGX agent

arXiv:2603.27577v2 Announce Type: replace Abstract: Vision-Language Navigation (VLN) requires an embodied agent to navigate complex environments by following natural language instructions, which typic

Structured Redundancy Modeling for Efficient Visual Token Pruning in High-Resolution MLLMs

SafetyDGX agent

arXiv:2607.23046v1 Announce Type: new Abstract: Recent high-resolution Multimodal Large Language Models (MLLMs) generate thousands of visual tokens per input, leading to a visual token explosion that

Subject-Level Heterogeneity in EEG Motor Imagery Decoding: A Large-Scale Benchmark and Portfolio-Based Reduction of the Search Space

Model ReleasesDGX agent

arXiv:2607.22778v1 Announce Type: cross Abstract: Robust EEG motor imagery decoding remains limited by strong inter-individual variability, making it difficult to identify pipelines that generalize ac

Success Is Not Self-Explanatory: Auditing Success Provenance in Agent Evaluation

Model ReleasesDGX agent

arXiv:2607.24054v1 Announce Type: new Abstract: A correct answer can conceal why an agent succeeded. Once agents change their information state during evaluation, correctness no longer distinguishes i

Super impressed with @huggingface's breakdown of the AI agent autonomous cyber attack from OpenAI, including technical timeline, & interacti…

AgentsDGX agent

Super impressed with @huggingface's breakdown of the AI agent autonomous cyber attack from OpenAI, including technical timeline, & interactive replay (AND how they defended against the attack)! Here i

Superpixel-Based QUBO for Scalable Quantum-Enhanced Medical Image Segmentation

ResearchDGX agent

arXiv:2607.24288v1 Announce Type: new Abstract: Quadratic unconstrained binary optimization (QUBO) has emerged as a powerful framework for medical computing problems. Binary decision variables natural

Surgical Re-enactment for Operating Room Workflow Datasets

TutorialsDGX agent

arXiv:2607.24206v1 Announce Type: cross Abstract: The introduction of new technologies, such as surgical robots, is driving the vision of a connected, smart operating room (OR). However, realizing thi

SWE-rebench Multilingual Update (Go, Java, Python, Rust, TS). Evaluated: GLM-5.2, DeepSeek-V4 Pro, Qwen3.6-27B and others

Model ReleasesDGX agent

Hi everyone! We’ve just released a major update to the leaderboard! We are expanding beyond Python with a new multilingual slice featuring real-world software engineering tasks across 5 languages. Ope

SwitchBraidNet: Quantisation-Aware Lightweight Architecture for Hybrid Brain-Computer Interface

ResearchDGX agent

arXiv:2606.18816v2 Announce Type: replace-cross Abstract: Hybrid brain-computer interfaces (BCIs) that integrate motor imagery (MI) and steady-state visual evoked potentials (SSVEP) provide high-dimen

SymStep: Symbolic Step Verification for Logical Reasoning

Model ReleasesDGX agent

arXiv:2607.23055v1 Announce Type: new Abstract: Chain-of-thought (CoT) prompting can fail severely on constraint-dense logical reasoning tasks, where unverified errors accumulate silently across steps

Synthetic Scenario Generation for Evaluation of Industry 4.0 Agents

SafetyDGX agent

arXiv:2607.22563v1 Announce Type: new Abstract: Industrial agent benchmarks require realistic evaluation scenarios that integrate telemetry, failure modes, maintenance records, and domain standards. H

SyRuP: Enhancing System-Prompt Following via Reward-Guided Prediction in LLM Decoding

SafetyDGX agent

arXiv:2607.23991v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly controlled through system prompts that specify roles, styles, formats, and safety requirements. However,

Systematic Analysis of Large Language Models and Transformer-Based Machine Translation for English-Tamil and Tamil-English Across Diverse Datasets

SafetyDGX agent

arXiv:2607.24515v1 Announce Type: new Abstract: The challenge of Machine Translation for low resource languages such as Tamil is primarily caused by the restricted amount of parallel data for these la

TableMind: An Autonomous Programmatic Agent for Tool-Augmented Table Reasoning

AgentsDGX agent

arXiv:2509.06278v4 Announce Type: replace Abstract: Table reasoning requires models to jointly perform comprehensive semantic understanding and precise numerical operations. Although recent large lang

Tag Questions and the Generational Reversal of Sycophancy Across 45 Language Models

Model ReleasesDGX agent

arXiv:2607.23976v1 Announce Type: cross Abstract: Appending a two-word confirmation tag to a decision question -- 'Is X the better choice?' versus 'X is the better choice, right?' -- changes whether a

Tailored untruths: How personalisation challenges LLM safeguards

SafetyDGX agent

arXiv:2510.12993v3 Announce Type: replace Abstract: Large Language Models (LLMs) can generate highly persuasive disinformation, yet little is known about how effectively they personalise it across lan

TaoMate: Anchor-Guided Memory Bridging Evolving and Reference States for Real-Time Audio-Video Digital Human Generation

Local AiDGX agent

arXiv:2607.24359v1 Announce Type: new Abstract: Real-time long-form digital-human generation relies on causal models to extend audio-visual content while preserving subject appearance and audio-video

Task-Conditional Faithfulness Auditing of Multimodal LLMs for Grid Diagnosis

ResearchDGX agent

arXiv:2607.24539v1 Announce Type: new Abstract: Multimodal large language models (LLMs) can combine topology, measurements, and incident text for grid diagnosis, yet answer accuracy does not establish

Teacher Knows It Best: Spontaneous Symmetry Breaking and Tipping Points in Networked Langevin Dynamics AI Sycophancy

ResearchDGX agent

arXiv:2607.24304v1 Announce Type: cross Abstract: We formulate a statistical physics framework to model a networked stochastic dynamical system exhibiting bistability, driven by additive noise and soc

TEmBed-T: A Multi-Dimensional Benchmark for Table-Level Embeddings

Model ReleasesDGX agent

arXiv:2607.24130v1 Announce Type: cross Abstract: Tabular data is the dominant structured-data modality, and learning table representations has become a core research direction. Table-level embeddings

← Previous
1…190191192193194…1411
Next →