AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
All
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,512 results
Model Releases

Scheming Ability in LLM-to-LLM Strategic Interactions

DGX agent

arXiv:2510.12826v2 Announce Type: replace-cross Abstract: As large language model (LLM) agents are deployed autonomously in diverse contexts, evaluating their capacity for strategic deception becomes

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

ScoringBench: A Benchmark for Evaluating Tabular Foundation Models with Proper Scoring Rules

DGX agent

arXiv:2603.29928v2 Announce Type: replace Abstract: Tabular foundation models such as TabPFN and TabICL already produce full predictive distributions, yet prevailing regression benchmarks evaluate the

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Secure On-Premise Deployment of Open-Weights Large Language Models in Radiology: An Isolation-First Architecture with Prospective Pilot Evaluation

DGX agent

arXiv:2604.22768v1 Announce Type: cross Abstract: Purpose: To design, implement, evaluate, and report on the regulatory requirements of a self-hosted LLM infrastructure for radiology adhering to the p

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

See Further, Think Deeper: Advancing VLM's Reasoning Ability with Low-level Visual Cues and Reflection

DGX agent

arXiv:2604.24339v1 Announce Type: cross Abstract: Recent advances in Vision-Language Models (VLMs) have benefited from Reinforcement Learning (RL) for enhanced reasoning. However, existing methods sti

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Seeing Is No Longer Believing: Frontier Image Generation Models, Synthetic Visual Evidence, and Real-World Risk

DGX agent

arXiv:2604.24197v1 Announce Type: cross Abstract: Frontier image generation has moved from artistic synthesis toward synthetic visual evidence. Systems such as GPT Image 2, Nano Banana Pro, Nano Banan

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Semantic Segmentation for Histopathology using Learned Regularization based on Global Proportions

DGX agent

arXiv:2604.24347v1 Announce Type: cross Abstract: In pathology, the spatial distribution and proportions of tissue types are key indicators of disease progression, and are more readily available than

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

SemiGDA: Generative Dual-distribution Alignment for Semi-Supervised Medical Image Segmentation

DGX agent

arXiv:2604.23274v1 Announce Type: new Abstract: Semi-supervised learning addresses label scarcity and high annotation costs in medical image segmentation by exploiting the latent information in unlabe

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

ServImage: An Image Generation and Editing Benchmark from Real-world Commercial Imaging Services

DGX agent

arXiv:2604.24023v1 Announce Type: new Abstract: Recent image generation and editing models demonstrate robust adherence to instructions and high visual quality on academic benchmarks. However, their p

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

SEVerA: Verified Synthesis of Self-Evolving Agents

DGX agent

arXiv:2603.25111v2 Announce Type: replace Abstract: Recent advances have shown the effectiveness of self-evolving LLM agents on tasks such as program repair and scientific discovery. In this paradigm,

model-releasesarxiv-cs-lg
28 Apr 2026
Model Releases

SFT-then-RL Outperforms Mixed-Policy Methods for LLM Reasoning

DGX agent

arXiv:2604.23747v1 Announce Type: cross Abstract: Recent mixed-policy optimization methods for LLM reasoning that interleave or blend supervised and reinforcement learning signals report improvements

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Shape: A Self-Supervised 3D Geometry Foundation Model for Industrial CAD Analysis

DGX agent

arXiv:2604.22826v1 Announce Type: new Abstract: Industrial CAD workflows require robust, generalizable 3D geometric representations supporting accuracy and explainability. We introduce Shape, a self-s

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

ShredBench: Evaluating the Semantic Reasoning Capabilities of Multimodal LLMs in Document Reconstruction

DGX agent

arXiv:2604.23813v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable performance in Visually Rich Document Understanding (VRDU) tasks, but their capabili

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

SIV-Bench: A Video Benchmark for Social Interaction Understanding and Reasoning

DGX agent

arXiv:2506.05425v2 Announce Type: replace-cross Abstract: Understanding social interaction, which encompasses perceiving numerous and subtle multimodal cues, inferring unobservable mental states and r

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

SketchVLM: Vision language models can annotate images to explain thoughts and guide users

DGX agent

arXiv:2604.22875v1 Announce Type: cross Abstract: When answering questions about images, humans naturally point, label, and draw to explain their reasoning. In contrast, modern vision-language models

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Skill Retrieval Augmentation for Agentic AI

DGX agent

arXiv:2604.24594v1 Announce Type: cross Abstract: As large language models (LLMs) evolve into agentic problem solvers, they increasingly rely on external, reusable skills to handle tasks beyond their

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

SMSI: System Model Security Inference: Automated Threat Modeling for Cyber-Physical Systems

DGX agent

arXiv:2604.23905v1 Announce Type: cross Abstract: Threat modeling for cyber-physical systems (CPS) remains a largely manual exercise. This project presents SMSI (System Model Security Inference), a hy

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Snap launches AI Sponsored Snaps, a conversational ad format in Snapchat's Chat tab that lets users talk to brand-specific AI agents for product recommendations (Aisha Malik/TechCrunch)

DGX agent

Aisha Malik / TechCrunch: Snap launches AI Sponsored Snaps, a conversational ad format in Snapchat's Chat tab that lets users talk to brand-specific AI agents for product recommendations — Snapchat an

model-releasestechmeme
28 Apr 2026
Model Releases

SoccerRef-Agents: Multi-Agent System for Automated Soccer Refereeing

DGX agent

arXiv:2604.23392v1 Announce Type: new Abstract: Refereeing is vital in sports, where fair, accurate, and explainable decisions are fundamental. While intelligent assistant technologies are being widel

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

SolarFCD: A Large-Scale Dataset and Benchmark for Solar Fault Classification in Photovoltaic Systems

DGX agent

arXiv:2604.23662v1 Announce Type: new Abstract: The increasing global deployment of solar photovoltaic (PV) systems needs robust, scalable, and automated inspection technologies capable of detecting a

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

SPAGS: Sparse-View Articulated Object Reconstruction from Single State via Planar Gaussian Splatting

DGX agent

arXiv:2511.17092v4 Announce Type: replace Abstract: Articulated objects are ubiquitous in daily environments, and their 3D reconstruction holds great significance across various fields. However, exist

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

Spatiotemporal Degradation-Aware 3D Gaussian Splatting for Realistic Underwater Scene Reconstruction

DGX agent

arXiv:2604.23551v1 Announce Type: new Abstract: Reconstructing realistic underwater scenes from underwater video remains a meaningful yet challenging task in the multimedia domain. The inherent spatio

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

SpecRLBench: A Benchmark for Generalization in Specification-Guided Reinforcement Learning

DGX agent

arXiv:2604.24729v1 Announce Type: new Abstract: Specification-guided reinforcement learning (RL) provides a principled framework for encoding complex, temporally extended tasks using formal specificat

model-releasesarxiv-cs-lg
28 Apr 2026
Model Releases

Speech Enhancement Based on Drifting Models

DGX agent

arXiv:2604.24199v1 Announce Type: cross Abstract: We propose Speech Enhancement based on Drifting Models (DriftSE), a novel generative framework that formulates denoising as an equilibrium problem. Ra

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Speech-FT: Merging Pre-trained And Fine-Tuned Speech Representation Models For Cross-Task Generalization

DGX agent

arXiv:2502.12672v4 Announce Type: replace-cross Abstract: Fine-tuning speech representation models can enhance performance on specific tasks but often compromises their cross-task generalization abili

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Sphere-Depth: A Benchmark for Depth Estimation Methods with Varying Spherical Camera Orientations

DGX agent

arXiv:2604.23432v1 Announce Type: cross Abstract: Reliable depth estimation from spherical images is crucial for 360{eg} vision in robotic navigation and immersive scene understanding. However, the on

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

SSR-Zero: Simple Self-Rewarding Reinforcement Learning for Machine Translation

DGX agent

arXiv:2505.16637v4 Announce Type: replace-cross Abstract: Large language models (LLMs) have recently demonstrated remarkable capabilities in machine translation (MT). However, most advanced MT-specifi

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Stabilizing Efficient Reasoning with Step-Level Advantage Selection

DGX agent

arXiv:2604.24003v1 Announce Type: new Abstract: Large language models (LLMs) achieve strong reasoning performance by allocating substantial computation at inference time, often generating long and ver

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

StoryTR: Narrative-Centric Video Temporal Retrieval with Theory of Mind Reasoning

DGX agent

arXiv:2604.23198v1 Announce Type: new Abstract: Current video moment retrieval excels at action-centric tasks but struggles with narrative content. Models can see extit{what is happening} but fail to

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Strategic Bidding in 6G Spectrum Auctions with Large Language Models

DGX agent

arXiv:2604.24156v1 Announce Type: cross Abstract: Efficient and fair spectrum allocation is a central challenge in 6G networks, where massive connectivity and heterogeneous services continuously compe

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

StratRAG: A Multi-Hop Retrieval Evaluation Dataset for Retrieval-Augmented Generation Systems

DGX agent

arXiv:2604.22757v1 Announce Type: cross Abstract: We introduce StratRAG, an open-source retrieval evaluation dataset for benchmarking Retrieval-Augmented Generation (RAG) systems on multi-hop reasonin

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Stress-Testing Emotional Support Models: Moving from Homogeneous to Diverse Help Seekers

DGX agent

arXiv:2601.07698v2 Announce Type: replace Abstract: As emotional support chatbots have recently gained significant traction across both research and industry, a common evaluation strategy has emerged:

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval

DGX agent

arXiv:2601.20597v2 Announce Type: replace Abstract: Continual Text-to-Video Retrieval (CTVR) is a challenging multimodal continual learning setting, where models must incrementally learn new semantic

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

Structural Pruning of Large Vision Language Models: A Comprehensive Study on Pruning Dynamics, Recovery, and Data Efficiency

DGX agent

arXiv:2604.24380v1 Announce Type: new Abstract: While Large Vision Language Models (LVLMs) demonstrate impressive capabilities, their substantial computational and memory requirements pose deployment

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Supernodes and Halos: Loss-Critical Hubs in LLM Feed-Forward Layers

DGX agent

arXiv:2604.23475v1 Announce Type: cross Abstract: We study the organization of channel-level importance in transformer feed-forward networks (FFNs). Using a Fisher-style loss proxy (LP) based on activ

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

SWE-QA: Can Language Models Answer Repository-level Code Questions?

DGX agent

arXiv:2509.14635v2 Announce Type: replace Abstract: Understanding and reasoning about entire software repositories is an essential capability for intelligent software engineering tools. While existing

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

SwissGov-RSD: A Human-annotated, Cross-lingual Benchmark for Token-level Recognition of Semantic Differences Between Related Documents

DGX agent

arXiv:2512.07538v3 Announce Type: replace Abstract: Recognizing semantic differences across documents is crucial for text generation evaluation and content alignment, especially in cross-lingual setti

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Switch Attention: Towards Dynamic and Fine-grained Hybrid Transformers

DGX agent

arXiv:2603.26380v2 Announce Type: replace Abstract: The attention mechanism has been the core component in modern transformer architectures. However, the computation of standard full attention scales

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

SycoPhantasy: Quantifying Sycophancy and Hallucination in Small Open Weight VLMs for Vision-Language Scoring of Fantasy Characters

DGX agent

arXiv:2604.24346v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly deployed as evaluators in tasks requiring nuanced image understanding, yet their reliability in scoring

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Symmetric Equilibrium Propagation for Thermodynamic Diffusion Training

DGX agent

arXiv:2604.23806v1 Announce Type: cross Abstract: The reverse process in score-based diffusion models is formally equivalent to overdamped Langevin dynamics in a time-dependent energy landscape. In ou

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

SynthPert: Enhancing LLM Biological Reasoning via Synthetic Reasoning Traces for Cellular Perturbation Prediction

DGX agent

arXiv:2509.25346v2 Announce Type: replace Abstract: Predicting cellular responses to genetic perturbations represents a fundamental challenge in systems biology, critical for advancing therapeutic dis

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

TACO: Efficient Communication Compression of Intermediate Tensors for Scalable Tensor-Parallel LLM Training

DGX agent

arXiv:2604.24088v1 Announce Type: cross Abstract: Handling communication overhead in large-scale tensor-parallel training remains a critical challenge due to the dense, near-zero distributions of inte

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Test of Time: Rethinking Temporal Signal of Benchmark Contamination

DGX agent

arXiv:2509.00072v3 Announce Type: replace Abstract: Post-cutoff performance decay has been widely interpreted as a temporal signal for benchmark contamination. We critically examine this belief and de

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

TexOCR: Advancing Document OCR Models for Compilable Page-to-LaTeX Reconstruction

DGX agent

arXiv:2604.22880v1 Announce Type: new Abstract: Existing document OCR largely targets plain text or Markdown, discarding the structural and executable properties that make LaTeX essential for scientif

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

TextGround4M: A Prompt-Aligned Dataset for Layout-Aware Text Rendering

DGX agent

arXiv:2604.24459v1 Announce Type: new Abstract: Despite recent advances in text-to-image generation, models still struggle to accurately render prompt-specified text with correct spatial layout -- esp

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

The Alignment Target Problem: Divergent Moral Judgments of Humans, AI Systems, and Their Designers

DGX agent

arXiv:2604.24155v1 Announce Type: cross Abstract: The quest to align machine behavior with human values raises fundamental questions about the moral frameworks that should govern AI decision-making. M

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

The California Coastal Commission has issued a formal apology to @elonmusk and SpaceX, adding that it will not consider political views or s…

DGX agent

The California Coastal Commission has issued a formal apology to @elonmusk and SpaceX, adding that it will not consider political views or speech in future regulatory decisions. • The Commission admit

model-releaseselon-musk--x
28 Apr 2026
Model Releases

The cracks inside OpenAI are deepening, and the numbers don’t lie. When your own CFO is sounding the alarm, something is seriously wrong. Ch…

DGX agent

The cracks inside OpenAI are deepening, and the numbers don’t lie. When your own CFO is sounding the alarm, something is seriously wrong. Check this out: 1: OpenAI missed its target of 1 billion weekl

model-releasesgary-marcus--x
28 Apr 2026
Model Releases

The FIDO Alliance launches two working groups to establish industry standards for securing AI agent transactions; Google contributes the Agent Payments Protocol (Lily Hay Newman/Wired)

DGX agent

Lily Hay Newman / Wired: The FIDO Alliance launches two working groups to establish industry standards for securing AI agent transactions; Google contributes the Agent Payments Protocol — AI agents ma

model-releasestechmeme
28 Apr 2026
← Previous
1…385386387388389…469
Next →