AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,507 results
28 Apr 2026

Qdrant Cloud launches high-performance vector database features for AI workloads

Model ReleasesDGX agent

Open-source vector database startup Qdrant Solutions GmbH today announced three new enterprise-grade capabilities on its cloud service to address performance, availability and compliance requirements

QED: An Open-Source Multi-Agent System for Generating Mathematical Proofs on Open Problems

Model ReleasesDGX agent

arXiv:2604.24021v1 Announce Type: new Abstract: We explore a central question in AI for mathematics: can AI systems produce original, nontrivial proofs for open research problems? Despite strong bench

QEVA: A Reference-Free Evaluation Metric for Narrative Video Summarization with Multimodal Question Answering

Model Releases

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2604.24052v1 Announce Type: cross Abstract: Video-to-text summarization remains underexplored in terms of comprehensive evaluation methods. Traditional n-gram overlap-based metrics and recent la

Quantum Knowledge Graph: Modeling Context-Dependent Triplet Validity

Model ReleasesDGX agent

arXiv:2604.23972v1 Announce Type: cross Abstract: Knowledge graphs (KGs) are increasingly used to support large lan guage model (LLM) reasoning, but standard triplet-based KGs treat each relation as g

Quasi-Equivariant Metanetworks

Model ReleasesDGX agent

arXiv:2604.23720v1 Announce Type: new Abstract: Metanetworks are neural architectures designed to operate directly on pretrained weights to perform downstream tasks. However, the parameter space serve

Quoting OpenAI Codex base_instructions

Model ReleasesDGX agent

Never talk about goblins, gremlins, raccoons, trolls, ogres, pigeons, or other animals or creatures unless it is absolutely and unambiguously relevant to the user's query. — OpenAI Codex base_instruct

RACANet: Reliability-Aware Crowd Anchor Network for RGB-T Crowd Counting

Model ReleasesDGX agent

arXiv:2604.24543v1 Announce Type: new Abstract: RGB-Thermal (T) crowd counting aims to integrate visible-spectrum and thermal infrared information to improve the robustness of crowd density estimation

Rank, Head-Channel Non-Identifiability, and Symmetry Breaking: A Precise Analysis of Representational Collapse in Transformers

Model ReleasesDGX agent

arXiv:2604.23681v1 Announce Type: cross Abstract: A widely cited result by Dong et al. (2021) showed that Transformers built from self-attention alone, without skip connections or feed-forward layers,

RAS: a Reliability Oriented Metric for Automatic Speech Recognition

Model ReleasesDGX agent

arXiv:2604.24278v1 Announce Type: cross Abstract: Automatic speech recognition systems often produce confident yet incorrect transcriptions under noisy or ambiguous conditions, which can be misleading

RAT: RunAnyThing via Fully Automated Environment Configuration

Model ReleasesDGX agent

arXiv:2604.23190v1 Announce Type: cross Abstract: Automating repository-level software engineering tasks is a foundational challenge for autonomous code agents, largely due to the difficulty of config

RaV-IDP: A Reconstruction-as-Validation Framework for Faithful Intelligent Document Processing

Model ReleasesDGX agent

arXiv:2604.23644v1 Announce Type: cross Abstract: Intelligent document processing pipelines extract structured entities (tables, images, and text) from documents for use in downstream systems such as

RCSB PDB AI Help Desk: retrieval-augmented generation for protein structure deposition support

Model ReleasesDGX agent

arXiv:2604.22800v1 Announce Type: cross Abstract: Motivation: Structural Biologists have contributed more than 245,000 experimentally determined three-dimensional structures of biological macromolecul

Read the blog: https://www.together.ai/blog/together-ai-brings-nvidia-nemotron-3-nano-omni-to-developers-on-day-0#

Model ReleasesDGX agent

Together AI announced the availability of NVIDIA Nemotron-3 Nano Omni models to developers on day zero of release, enabling early access to these multimodal AI models through their platform. The annou

Reading in the Dark: Low-light Scene Text Recognition

Model ReleasesDGX agent

arXiv:2604.23685v1 Announce Type: new Abstract: Accurate text recognition in low-light environments is essential for intelligent systems in applications ranging from autonomous vehicles to smart surve

Real-time windrow detection from onboard tractor sensors for automated following

Model ReleasesDGX agent

arXiv:2604.24628v1 Announce Type: new Abstract: Proprietary design in commercial windrow-detection systems restricts transparency and limits progress in open autonomous forage-harvesting research. We

RealFin: How Well Do LLMs Reason About Finance When Users Leave Things Unsaid?

Model ReleasesDGX agent

arXiv:2602.07096v2 Announce Type: replace-cross Abstract: Reliable financial reasoning requires knowing not only how to answer, but also when an answer cannot be justified. In real financial practice,

Reasonably reasoning AI agents can avoid game-theoretic failures in zero-shot, provably

Model ReleasesDGX agent

arXiv:2603.18563v2 Announce Type: replace Abstract: As autonomous AI agents increasingly mediate online platform markets, a fundamental question emerges: do these markets generate stable strategic out

Reclaiming Residual Knowledge: A Novel Paradigm to Low-Bit Quantization

Model ReleasesDGX agent

arXiv:2408.00923v2 Announce Type: replace-cross Abstract: This paper explores a novel paradigm in low-bit (i.e. 4-bits or lower) quantization, differing from existing state-of-the-art methods, by fram

RefEvo: Agentic Design with Co-Evolutionary Verification for Agile Reference Model Generation

Model ReleasesDGX agent

arXiv:2604.24218v1 Announce Type: cross Abstract: As the complexity of System-on-Chip (SoC) designs grows, the shift-left paradigm necessitates the rapid development of high-fidelity reference models

Reinforcement Learning with Backtracking Feedback

Model ReleasesDGX agent

arXiv:2602.08377v2 Announce Type: replace-cross Abstract: Addressing the critical need for robust safety in Large Language Models (LLMs), particularly against adversarial attacks and in-distribution e

ResAF-Net: An Anchor-Free Attention-Based Network for Tree Detection and Agricultural Mapping in Palestine

Model ReleasesDGX agent

arXiv:2604.23653v1 Announce Type: cross Abstract: Reliable agricultural data is essential for food security, land-use planning, and economic resilience, yet in Palestine, such data remains difficult t

Resolution scaling governs DINOv3 transfer performance in chest radiograph classification

Model ReleasesDGX agent

arXiv:2510.07191v3 Announce Type: replace-cross Abstract: Self-supervised learning (SSL) has improved visual representation learning, but its value in chest radiography remains uncertain. DINOv3 exten

Resource-Lean Lexicon Induction for German Dialects

Model ReleasesDGX agent

arXiv:2604.23824v1 Announce Type: new Abstract: Automatic induction of high-quality dictionaries is essential for building lexical resources, yet low-resource languages and dialects pose several chall

Results on the remainder of our benchmarks are in for Kimi K2.6, the new #1 open-weight model on the Vals Index (#8 overall) Kimi K2.6 is co…

Model ReleasesDGX agent

Results on the remainder of our benchmarks are in for Kimi K2.6, the new #1 open-weight model on the Vals Index (#8 overall) Kimi K2.6 is competitive with many closed-weight models, at a fraction of t

Rethinking Parameter Sharing for LLM Fine-Tuning with Multiple LoRAs

Model ReleasesDGX agent

arXiv:2509.25414v2 Announce Type: replace-cross Abstract: Large language models are often adapted using parameter-efficient techniques such as Low-Rank Adaptation (LoRA), formulated as y = W_0x + BAx,

ReVSI: Rebuilding Visual Spatial Intelligence Evaluation for Accurate Assessment of VLM 3D Reasoning

Model ReleasesDGX agent

arXiv:2604.24300v1 Announce Type: new Abstract: Current evaluations of spatial intelligence can be systematically invalid under modern vision-language model (VLM) settings. First, many benchmarks deri

Robust Audio-Text Retrieval via Cross-Modal Attention and Hybrid Loss

Model ReleasesDGX agent

arXiv:2604.23323v1 Announce Type: new Abstract: Audio-text retrieval enables semantic alignment between audio content and natural language queries, supporting applications in multimedia search, access

RouteNLP: Closed-Loop LLM Routing with Conformal Cascading and Distillation Co-Optimization

Model ReleasesDGX agent

arXiv:2604.23577v1 Announce Type: new Abstract: Serving diverse NLP workloads with large language models is costly: at one enterprise partner, inference costs exceeded $200K/month despite over 70% of

Scalable Agentic Reasoning for Designing Biologics Targeting Intrinsically Disordered Proteins

Model ReleasesDGX agent

arXiv:2512.15930v2 Announce Type: replace-cross Abstract: Intrinsically disordered proteins (IDPs) represent crucial therapeutic targets due to their significant role in disease -- approximately 80% o

Scaling Multi-Node Mixture-of-Experts Inference Using Expert Activation Patterns

Model ReleasesDGX agent

arXiv:2604.23150v1 Announce Type: cross Abstract: Most recent state-of-the-art (SOTA) large language models (LLMs) use Mixture-of-Experts (MoE) architectures to scale model capacity without proportion

Scaling Properties of Continuous Diffusion Spoken Language Models

Model ReleasesDGX agent

arXiv:2604.24416v1 Announce Type: cross Abstract: Speech-only spoken language models (SLMs) lag behind text and text-speech models in performance, with recent discrete autoregressive (AR) SLMs indicat

Scheming Ability in LLM-to-LLM Strategic Interactions

Model ReleasesDGX agent

arXiv:2510.12826v2 Announce Type: replace-cross Abstract: As large language model (LLM) agents are deployed autonomously in diverse contexts, evaluating their capacity for strategic deception becomes

ScoringBench: A Benchmark for Evaluating Tabular Foundation Models with Proper Scoring Rules

Model ReleasesDGX agent

arXiv:2603.29928v2 Announce Type: replace Abstract: Tabular foundation models such as TabPFN and TabICL already produce full predictive distributions, yet prevailing regression benchmarks evaluate the

Secure On-Premise Deployment of Open-Weights Large Language Models in Radiology: An Isolation-First Architecture with Prospective Pilot Evaluation

Model ReleasesDGX agent

arXiv:2604.22768v1 Announce Type: cross Abstract: Purpose: To design, implement, evaluate, and report on the regulatory requirements of a self-hosted LLM infrastructure for radiology adhering to the p

See Further, Think Deeper: Advancing VLM's Reasoning Ability with Low-level Visual Cues and Reflection

Model ReleasesDGX agent

arXiv:2604.24339v1 Announce Type: cross Abstract: Recent advances in Vision-Language Models (VLMs) have benefited from Reinforcement Learning (RL) for enhanced reasoning. However, existing methods sti

Seeing Is No Longer Believing: Frontier Image Generation Models, Synthetic Visual Evidence, and Real-World Risk

Model ReleasesDGX agent

arXiv:2604.24197v1 Announce Type: cross Abstract: Frontier image generation has moved from artistic synthesis toward synthetic visual evidence. Systems such as GPT Image 2, Nano Banana Pro, Nano Banan

Semantic Segmentation for Histopathology using Learned Regularization based on Global Proportions

Model ReleasesDGX agent

arXiv:2604.24347v1 Announce Type: cross Abstract: In pathology, the spatial distribution and proportions of tissue types are key indicators of disease progression, and are more readily available than

SemiGDA: Generative Dual-distribution Alignment for Semi-Supervised Medical Image Segmentation

Model ReleasesDGX agent

arXiv:2604.23274v1 Announce Type: new Abstract: Semi-supervised learning addresses label scarcity and high annotation costs in medical image segmentation by exploiting the latent information in unlabe

ServImage: An Image Generation and Editing Benchmark from Real-world Commercial Imaging Services

Model ReleasesDGX agent

arXiv:2604.24023v1 Announce Type: new Abstract: Recent image generation and editing models demonstrate robust adherence to instructions and high visual quality on academic benchmarks. However, their p

SEVerA: Verified Synthesis of Self-Evolving Agents

Model ReleasesDGX agent

arXiv:2603.25111v2 Announce Type: replace Abstract: Recent advances have shown the effectiveness of self-evolving LLM agents on tasks such as program repair and scientific discovery. In this paradigm,

SFT-then-RL Outperforms Mixed-Policy Methods for LLM Reasoning

Model ReleasesDGX agent

arXiv:2604.23747v1 Announce Type: cross Abstract: Recent mixed-policy optimization methods for LLM reasoning that interleave or blend supervised and reinforcement learning signals report improvements

Shape: A Self-Supervised 3D Geometry Foundation Model for Industrial CAD Analysis

Model ReleasesDGX agent

arXiv:2604.22826v1 Announce Type: new Abstract: Industrial CAD workflows require robust, generalizable 3D geometric representations supporting accuracy and explainability. We introduce Shape, a self-s

ShredBench: Evaluating the Semantic Reasoning Capabilities of Multimodal LLMs in Document Reconstruction

Model ReleasesDGX agent

arXiv:2604.23813v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable performance in Visually Rich Document Understanding (VRDU) tasks, but their capabili

SIV-Bench: A Video Benchmark for Social Interaction Understanding and Reasoning

Model ReleasesDGX agent

arXiv:2506.05425v2 Announce Type: replace-cross Abstract: Understanding social interaction, which encompasses perceiving numerous and subtle multimodal cues, inferring unobservable mental states and r

SketchVLM: Vision language models can annotate images to explain thoughts and guide users

Model ReleasesDGX agent

arXiv:2604.22875v1 Announce Type: cross Abstract: When answering questions about images, humans naturally point, label, and draw to explain their reasoning. In contrast, modern vision-language models

Skill Retrieval Augmentation for Agentic AI

Model ReleasesDGX agent

arXiv:2604.24594v1 Announce Type: cross Abstract: As large language models (LLMs) evolve into agentic problem solvers, they increasingly rely on external, reusable skills to handle tasks beyond their

SMSI: System Model Security Inference: Automated Threat Modeling for Cyber-Physical Systems

Model ReleasesDGX agent

arXiv:2604.23905v1 Announce Type: cross Abstract: Threat modeling for cyber-physical systems (CPS) remains a largely manual exercise. This project presents SMSI (System Model Security Inference), a hy

Snap launches AI Sponsored Snaps, a conversational ad format in Snapchat's Chat tab that lets users talk to brand-specific AI agents for product recommendations (Aisha Malik/TechCrunch)

Model ReleasesDGX agent

Aisha Malik / TechCrunch: Snap launches AI Sponsored Snaps, a conversational ad format in Snapchat's Chat tab that lets users talk to brand-specific AI agents for product recommendations — Snapchat an

SoccerRef-Agents: Multi-Agent System for Automated Soccer Refereeing

Model ReleasesDGX agent

arXiv:2604.23392v1 Announce Type: new Abstract: Refereeing is vital in sports, where fair, accurate, and explainable decisions are fundamental. While intelligent assistant technologies are being widel

SolarFCD: A Large-Scale Dataset and Benchmark for Solar Fault Classification in Photovoltaic Systems

Model ReleasesDGX agent

arXiv:2604.23662v1 Announce Type: new Abstract: The increasing global deployment of solar photovoltaic (PV) systems needs robust, scalable, and automated inspection technologies capable of detecting a

SPAGS: Sparse-View Articulated Object Reconstruction from Single State via Planar Gaussian Splatting

Model ReleasesDGX agent

arXiv:2511.17092v4 Announce Type: replace Abstract: Articulated objects are ubiquitous in daily environments, and their 3D reconstruction holds great significance across various fields. However, exist

Spatiotemporal Degradation-Aware 3D Gaussian Splatting for Realistic Underwater Scene Reconstruction

Model ReleasesDGX agent

arXiv:2604.23551v1 Announce Type: new Abstract: Reconstructing realistic underwater scenes from underwater video remains a meaningful yet challenging task in the multimedia domain. The inherent spatio

SpecRLBench: A Benchmark for Generalization in Specification-Guided Reinforcement Learning

Model ReleasesDGX agent

arXiv:2604.24729v1 Announce Type: new Abstract: Specification-guided reinforcement learning (RL) provides a principled framework for encoding complex, temporally extended tasks using formal specificat

Speech Enhancement Based on Drifting Models

Model ReleasesDGX agent

arXiv:2604.24199v1 Announce Type: cross Abstract: We propose Speech Enhancement based on Drifting Models (DriftSE), a novel generative framework that formulates denoising as an equilibrium problem. Ra

Speech-FT: Merging Pre-trained And Fine-Tuned Speech Representation Models For Cross-Task Generalization

Model ReleasesDGX agent

arXiv:2502.12672v4 Announce Type: replace-cross Abstract: Fine-tuning speech representation models can enhance performance on specific tasks but often compromises their cross-task generalization abili

Sphere-Depth: A Benchmark for Depth Estimation Methods with Varying Spherical Camera Orientations

Model ReleasesDGX agent

arXiv:2604.23432v1 Announce Type: cross Abstract: Reliable depth estimation from spherical images is crucial for 360{eg} vision in robotic navigation and immersive scene understanding. However, the on

SSR-Zero: Simple Self-Rewarding Reinforcement Learning for Machine Translation

Model ReleasesDGX agent

arXiv:2505.16637v4 Announce Type: replace-cross Abstract: Large language models (LLMs) have recently demonstrated remarkable capabilities in machine translation (MT). However, most advanced MT-specifi

Stabilizing Efficient Reasoning with Step-Level Advantage Selection

Model ReleasesDGX agent

arXiv:2604.24003v1 Announce Type: new Abstract: Large language models (LLMs) achieve strong reasoning performance by allocating substantial computation at inference time, often generating long and ver

StoryTR: Narrative-Centric Video Temporal Retrieval with Theory of Mind Reasoning

Model ReleasesDGX agent

arXiv:2604.23198v1 Announce Type: new Abstract: Current video moment retrieval excels at action-centric tasks but struggles with narrative content. Models can see extit{what is happening} but fail to

Strategic Bidding in 6G Spectrum Auctions with Large Language Models

Model ReleasesDGX agent

arXiv:2604.24156v1 Announce Type: cross Abstract: Efficient and fair spectrum allocation is a central challenge in 6G networks, where massive connectivity and heterogeneous services continuously compe

← Previous
1…308309310311312…376
Next →