AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,603 results
Model Releases

ShareChat: A Dataset of Chatbot Conversations in the Wild

DGX agent

arXiv:2512.17843v4 Announce Type: replace-cross Abstract: By evaluating Large Language Models (LLMs) through uniform, text-only interfaces, current academic benchmarks obscure how the unique designs a

model-releasesarxiv-cs-ai
19 May 2026
Model Releases
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

SIPO: Stabilized and Improved Preference Optimization for Aligning Diffusion Models

DGX agent

arXiv:2505.21893v3 Announce Type: replace-cross Abstract: Preference learning has garnered extensive attention as an effective technique for aligning diffusion models with human preferences in visual

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

SIREM: Speech-Informed MRI Reconstruction with Learned Sampling

DGX agent

arXiv:2605.18221v1 Announce Type: cross Abstract: Real-time magnetic resonance imaging (rtMRI) of speech production enables non-invasive visualization of dynamic vocal-tract motion and is valuable for

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

SkillGenBench: Benchmarking Skill Generation Pipelines for LLM Agents

DGX agent

arXiv:2605.18693v1 Announce Type: new Abstract: As LLM agents are increasingly built around reusable skills, a central challenge is no longer only whether agents can use provided skills, but whether t

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Skills on the Fly: Test-Time Adaptive Skill Synthesis for LLM Agents

DGX agent

arXiv:2605.16986v1 Announce Type: cross Abstract: LLM agents benefit from reusable skills, yet test-time tasks often require guidance more specific than a static skill library can provide. We propose

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

SkillsVote: Lifecycle Governance of Agent Skills from Collection, Recommendation to Evolution

DGX agent

arXiv:2605.18401v1 Announce Type: cross Abstract: Long-horizon LLM agents leave traces that could become reusable experience, but raw trajectories are noisy and hard to govern. We treat Agent Skills a

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

SkyNative: A Native Multimodal Framework for Remote Sensing Visual Evidence Reasoning

DGX agent

arXiv:2605.17949v1 Announce Type: new Abstract: Remote sensing vision-language models commonly rely on pretrained visual encoders to convert images into semantic features before language-model reasoni

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

SLEIGHT-Bench: A Benchmark of Evasion Attacks Against Agent Monitors

DGX agent

arXiv:2605.16626v1 Announce Type: cross Abstract: Since autonomous coding agents generate complex behaviors at high-volume, we may want to use other LLMs to monitor actions to reduce the risk from dan

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Small-scale photonic Kolmogorov-Arnold networks using standard telecom nonlinear modules

DGX agent

arXiv:2604.08432v2 Announce Type: replace-cross Abstract: Photonic neural networks promise ultrafast inference, yet most architectures rely on linear optical meshes with electronic nonlinearities, rei

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

So google is replacing gemini-cli with agy (antigravity cli), but: 1. agy is not opensource 2. It no longer supports ACP Really unfortunate …

DGX agent

Google is transitioning from the Gemini CLI tool to a new CLI called AGY (Antigravity CLI), but this change has drawbacks: AGY is not open source and no longer supports ACP functionality, representing

model-releasesjeremy-howard--x
19 May 2026
Model Releases

SocialMemBench: Are AI Memory Systems Ready for Social Group Settings?

DGX agent

arXiv:2605.17789v1 Announce Type: cross Abstract: Memory systems for AI assistants were built for single-user dialogue and fail characteristically when applied to multi-party social group settings. Th

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Sola Security launches Lumina to cut enterprise security alert noise with contextual AI

DGX agent

Cybersecurity startup Sola Security Ltd. today announced the launch of Lumina, an autonomous risk intelligence platform that applies contextual artificial intelligence across cloud, identity, software

model-releasessiliconangle
19 May 2026
Model Releases

SomaliWeb v1: A Quality-Filtered Somali Web Corpus with a Matched Tokenizer and a Public Language-Identification Benchmark

DGX agent

arXiv:2605.18232v1 Announce Type: cross Abstract: Somali is a Cushitic language of the Horn of Africa with ~25 million speakers, yet no documented dedicated Somali pretraining corpus with a companion

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Some fun Gemini Omni use cases from the community👇🧵 (We’ll keep updating this thread throughout the day)

DGX agent

This X thread from Google AI showcases practical and creative applications of Gemini Omni, Google's multimodal AI model, as demonstrated and shared by the user community. The thread appears to be a cu

model-releasesgoogle-ai--x
19 May 2026
Model Releases

Sometin Beta Pass Notin (SBPN): Improving Multilingual ASR for Nigerian Languages via Knowledge Distillation

DGX agent

arXiv:2605.17710v1 Announce Type: new Abstract: Although modern multilingual Automatic Speech Recognition (ASR) systems support several Nigerian languages, their performance consistently lags behind h

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Sparse Training of Neural Networks based on Multilevel Mirror Descent

DGX agent

arXiv:2602.03535v2 Announce Type: replace Abstract: We introduce a dynamic sparse training algorithm based on linearized Bregman iterations / mirror descent that exploits the naturally incurred sparsi

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

SPATIOROUTE: Dynamic Prompt Routing for Zero-Shot Spatial Reasoning

DGX agent

arXiv:2605.18209v1 Announce Type: cross Abstract: Spatial question answering over egocentric video is a challenging task that requires Vision-Language Models (VLMs) to reason about 3D object positions

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

SpecSem-Net: Integrating Spectral and Semantic Features for Robust AI-generated Video Detection

DGX agent

arXiv:2605.17311v1 Announce Type: new Abstract: The remarkable visual fidelity of recent commercial video generative models, such as Sora and Veo, renders robust AI-generated video detection increasin

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

Spherical VAE with Cluster-Aware Feasible Regions: Guaranteed Prevention of Posterior Collapse

DGX agent

arXiv:2603.10935v4 Announce Type: replace-cross Abstract: Variational autoencoders (VAEs) frequently suffer from posterior collapse, where the latent variables become uninformative as the approximate

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Spotify's Chief Architect just showed how they ship 4,5K deployments /day with Claude at Anthropic stage 27-minutes. free. By #1 music app d…

DGX agent

Spotify's Chief Architect just showed how they ship 4,5K deployments /day with Claude at Anthropic stage 27-minutes. free. By #1 music app dev 'More than 99% of our engineers use AI coding tools. Adop

model-releasesboris-cherny--x
19 May 2026
Model Releases

Stabilizing, Scaling & Enhancing MeanFlow for Large-scale Diffusion Distillation

DGX agent

arXiv:2605.17834v1 Announce Type: new Abstract: Diffusion models exhibit remarkable generative capability, but their high latency limits practical deployment. Many studies have attempted to reduce sam

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

StableVLA: Towards Robust Vision-Language-Action Models without Extra Data

DGX agent

arXiv:2605.18287v1 Announce Type: new Abstract: It is infeasible to encompass all possible disturbances within the training dataset. This raises a critical question regarding the robustness of Vision-

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

State-of-the-Art Claims Require State-of-the-Art Evidence

DGX agent

arXiv:2605.17273v1 Announce Type: cross Abstract: State-of-the-Art (SOTA) claims pervade Artificial Intelligence (AI) and Machine Learning (ML) research. These claims rest on benchmark evaluations, wh

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Statistical Limits and Efficient Algorithms for Differentially Private Federated Learning

DGX agent

arXiv:2605.18656v1 Announce Type: cross Abstract: Federated Learning is a leading framework for training ML and AI models collaboratively across numerous user devices or databases. We study the trade-

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

SteadyDancer: Harmonized and Coherent Human Image Animation with First-Frame Preservation

DGX agent

arXiv:2511.19320v2 Announce Type: replace Abstract: Preserving first-frame identity while ensuring precise motion control is a fundamental challenge in human image animation. The Image-to-Motion Bindi

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

Strategic Over-Parameterization for Generalizable Low-Rank Adaptation

DGX agent

arXiv:2605.16470v1 Announce Type: cross Abstract: Adapting large language models (LLMs) to downstream tasks via full fine-tuning is increasingly impractical due to its computational and memory demands

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

StreamPro: From Reactive Perception to Proactive Decision-Making in Streaming Video

DGX agent

arXiv:2605.16381v1 Announce Type: cross Abstract: Proactive streaming video understanding requires models to continuously process video streams and decide when to respond, rather than merely what to r

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

StrLoRA: Towards Streaming Continual Visual Instruction Tuning for MLLMs

DGX agent

arXiv:2605.16353v1 Announce Type: cross Abstract: Continual Visual Instruction Tuning (CVIT) enables Multimodal Large Language Models to incrementally acquire new abilities. However, existing CVIT met

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Structured Labeling Enables Faster Vision-Language Models for End-to-End Autonomous Driving

DGX agent

arXiv:2506.05442v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) offer a promising approach to end-to-end autonomous driving due to their human-like reasoning capabilities. Howe

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Structured Neural Marked Point Processes for Interpretable Event Interaction Modeling

DGX agent

arXiv:2605.17568v1 Announce Type: new Abstract: Multi-class event streams arise in numerous real-world applications, where uncovering structured, interpretable inter-event relationships, together with

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

STT-Arena: A More Realistic Environment for Tool-Using with Spatio-Temporal Dynamics

DGX agent

arXiv:2605.18548v1 Announce Type: cross Abstract: Large language models (LLMs) deployed in real-world agentic applications must be capable of replanning and adapting when mid-task disruptions invalida

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

StyleText: A Large-Scale Dataset and Benchmark for Stylized Scene Text Inpainting

DGX agent

arXiv:2605.17309v1 Announce Type: cross Abstract: We present StyleText, a large-scale dataset and benchmark for localized scene-text inpainting with style preservation. StyleText contains 28,518 image

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Supervise Less, See More: Training-free Nuclear Instance Segmentation with Prototype-Guided Prompting

DGX agent

arXiv:2511.19953v2 Announce Type: replace Abstract: Accurate nuclear instance segmentation is a pivotal task in computational pathology, supporting data-driven clinical insights and facilitating downs

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

Supervising the search process produces reliable and generalizable information-seeking agents

DGX agent

arXiv:2502.13957v3 Announce Type: replace-cross Abstract: Large language models (LLMs) are transforming web search by shifting from document ranking to synthesizing answers, and are increasingly deplo

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

SurgLQA: Scalable Long-Horizon Surgical Video Question Answering

DGX agent

arXiv:2605.17915v1 Announce Type: new Abstract: Surgical Video Question Answering (VideoQA) provides a promising paradigm for dynamic intraoperative interpretation, enabling real-time decision support

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

Sustainability via LLM Right-sizing

DGX agent

arXiv:2504.13217v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have become increasingly embedded in organizational workflows. This has raised concerns over their energy consump

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

SVFSearch: A Multimodal Knowledge-Intensive Benchmark for Short-Video Frame Search in the Gaming Vertical Domain

DGX agent

arXiv:2605.17946v1 Announce Type: new Abstract: Multimodal large language models are increasingly used as agent backbones that understand multimodal inputs, plan retrieval actions, invoke external too

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

SwordBench: Evaluating Orthogonality of Steering Image Representations

DGX agent

arXiv:2605.16372v1 Announce Type: cross Abstract: Steering or intervening on model representations at inference time to correct predictions is essential for AI interpretability and safety, yet existin

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Symmetry-Compatible Principle for Optimizer Design: Embeddings, LM Heads, SwiGLU MLPs, and MoE Routers

DGX agent

arXiv:2605.18106v1 Announce Type: cross Abstract: A striking geometric disparity has long persisted in the practice of deep learning. While modern neural network architectures naturally exhibit rich s

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Symphony for Speech-to-Text: Supporting Real-Time Medical Voice Interfaces

DGX agent

arXiv:2605.16545v1 Announce Type: cross Abstract: After decades of use in dictation and, more recently, ambient documentation, speech is emerging as a primary modality for interacting with technology

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

TabH2O: A Unified Foundation Model for Tabular Prediction

DGX agent

arXiv:2605.18383v1 Announce Type: new Abstract: We present TabH2O, a foundation model for tabular data that performs classification and regression in a single forward pass via in-context learning. Tab

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

Tactile-based Multimodal Fusion in Embodied Intelligence: A Survey of Vision, Language, and Contact-Driven Paradigms

DGX agent

arXiv:2605.17336v1 Announce Type: cross Abstract: Tactile sensing is a fundamental modality for embodied intelligence, offering unique and direct feedback on contact geometry, material properties, and

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

TailedTS: Benchmark Dataset for Heavy-Tailed Time Series Prediction and Periodicity Quantification

DGX agent

arXiv:2605.16361v1 Announce Type: cross Abstract: We present TailedTS, a large-scale benchmark dataset derived from Wikipedia hourly page view observations throughout 2024, specifically designed to te

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

TAME: Test-Time Adversarial Prompt Tuning via Mixture-of-Experts for Vision-Language Models

DGX agent

arXiv:2605.17577v1 Announce Type: new Abstract: Large-scale pre-trained Vision-Language models (VLMs), such as CLIP, exhibit strong zero-shot generalization, yet remain highly vulnerable to impercepti

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

Task Abstention for Large Language Models in Code Generation

DGX agent

arXiv:2605.17029v1 Announce Type: cross Abstract: Large language models (LLMs) have revolutionized automated code generation. One serious concern, however, is the so-called ``hallucination'', i.e., LL

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

TaskGround: Structured Executable Task Inference for Full-Scene Household Reasoning

DGX agent

arXiv:2605.18109v1 Announce Type: new Abstract: In real home deployments, household agents must often operate from a complete household scene and a situated household request, rather than from a clean

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Taxonomy and Consistency Analysis of Safety Benchmarks for AI Agents

DGX agent

arXiv:2605.16282v1 Announce Type: cross Abstract: The rapid deployment of LLM-based autonomous agents has introduced safety risks that extend far beyond traditional LLM concerns, prompting a prolifera

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

TeleCom-Bench: How Far Are Large Language Models from Industrial Telecommunication Applications?

DGX agent

arXiv:2605.18025v1 Announce Type: new Abstract: While Large Language Models have achieved remarkable integration in various vertical scenarios, their deployment in the telecommunications domain remain

model-releasesarxiv-cs-ai
19 May 2026
← Previous
1…300301302303304…471
Next →