AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

Revisiting Reinforcement Learning with Verifiable Rewards from a Contrastive Perspective

DGX agent

arXiv:2605.12969v1 Announce Type: cross Abstract: RLVR has become a widely adopted paradigm for improving LLMs' reasoning capabilities, and GRPO is one of its most representative algorithms. In this p

model-releasesarxiv-cs-ai
14 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

RISED: A Pre-Deployment Safety Evaluation Framework for Clinical AI Decision-Support Systems

DGX agent

arXiv:2605.12895v1 Announce Type: cross Abstract: Aggregate accuracy metrics dominate the evaluation of clinical AI decision-support systems but do not detect deployment-phase failures of input reliab

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

RoSplat: Robust Feed-Forward Pixel-wise Gaussian Splatting for Varying Input Views and High-Resolution Rendering

DGX agent

arXiv:2605.13093v1 Announce Type: new Abstract: Generalizable 3D Gaussian Splatting has recently emerged as an efficient approach for novel-view synthesis, enabling feed-forward synthesis from only a

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

RS-Claw: Progressive Active Tool Exploration via Hierarchical Skill Trees for Remote Sensing Agents

DGX agent

arXiv:2605.13391v1 Announce Type: new Abstract: The rise of multi-modal large language models (MLLMs) is shifting remote sensing (RS) intelligence from 'see' to 'action', as OpenClaw-style frameworks

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

RTLC -- Research, Teach-to-Learn, Critique: A three-stage prompting paradigm inspired by the Feynman Learning Technique that lifts LLM-as-judge accuracy on JudgeBench with no fine-tuning

DGX agent

arXiv:2605.13695v1 Announce Type: cross Abstract: LLM-as-a-judge is now the default measurement instrument for open-ended generation, but on the public JudgeBench benchmark even strong instruction-tun

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Safe Bayesian Optimization for Uncertain Correlations Matrices in Linear Models of Co-Regionalization

DGX agent

arXiv:2605.13302v1 Announce Type: new Abstract: This paper extends safety guarantees for multi-task Bayesian optimization with uncertain correlation matrices from intrinsic co-reginalization models to

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs

DGX agent

arXiv:2510.18245v3 Announce Type: replace-cross Abstract: Scaling the number of parameters and the size of training data has proven to be an effective strategy for improving large language model (LLM)

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

scShapeBench: Discovering geometry from high dimensional scRNAseq data

DGX agent

arXiv:2605.12662v1 Announce Type: new Abstract: High-dimensional point cloud data arise across many scientific domains, especially single-cell biology. The shapes or topologies of these datasets deter

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

SCU-Hand with Integrated Single-Sheet Valve: A Funnel-Shaped Robotic Hand for Milligram-Scale Powder Handling

DGX agent

arXiv:2512.07091v2 Announce Type: replace Abstract: Laboratory Automation (LA) has the potential to accelerate solid-state materials discovery by enabling continuous robotic operation without human in

model-releasesarxiv-cs-ro
14 May 2026
Model Releases

Seg-Agent: Test-Time Multimodal Reasoning for Training-Free Language-Guided Segmentation

DGX agent

arXiv:2605.12953v1 Announce Type: cross Abstract: Language-guided segmentation transcends the scope limitations of traditional semantic segmentation, enabling models to segment arbitrary target region

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Senses Wide Shut: A Representation-Action Gap in Omnimodal LLMs

DGX agent

arXiv:2605.13737v1 Announce Type: new Abstract: When an omnimodal large language model accepts a question whose textual premise contradicts what it actually sees or hears, does the failure lie in perc

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

SkySplat: Generalizable 3D Gaussian Splatting from Multi-Temporal Sparse Satellite Images

DGX agent

arXiv:2508.09479v2 Announce Type: replace Abstract: Three-dimensional scene reconstruction from sparse-view satellite images is a long-standing and challenging task. While 3D Gaussian Splatting (3DGS)

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

SMA: Submodular Modality Aligner For Data Efficient Multimodal Learning

DGX agent

arXiv:2605.12872v1 Announce Type: new Abstract: Despite the recent success of Multimodal Foundation Models (FMs), their reliance on massive paired datasets limits their applicability in low-data and r

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

Small Area Estimation of Case Growths for Timely COVID-19 Outbreak Detection

DGX agent

arXiv:2312.04110v2 Announce Type: replace-cross Abstract: The COVID-19 pandemic has exerted a profound impact on the global economy and continues to exact a significant toll on human lives. The COVID-

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

Sockpuppetting: Jailbreaking LLMs by Combining Prefilling with Optimization

DGX agent

arXiv:2601.13359v2 Announce Type: replace-cross Abstract: Prefill attacks are an effective and low-cost jailbreaking method, as they directly insert an acceptance sequence (e.g., 'Sure, here is how to

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

SpatialReward: Bridging the Perception Gap in Online RL for Image Editing via Explicit Spatial Reasoning

DGX agent

arXiv:2602.07458v4 Announce Type: replace Abstract: Online Reinforcement Learning (RL) offers a promising avenue for complex image editing but is currently constrained by the scarcity of reliable and

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

SpectralTrain: A Universal Framework for Hyperspectral Image Classification

DGX agent

arXiv:2511.16084v2 Announce Type: replace-cross Abstract: Hyperspectral image (HSI) classification typically involves large-scale data and computationally intensive training, which limits the practica

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

SpikeProphecy: A Large-Scale Benchmark for Autoregressive Neural Population Forecasting

DGX agent

arXiv:2605.12992v1 Announce Type: cross Abstract: Neural population models, which predict the joint firing of many simultaneously recorded neurons forward in time, are typically evaluated by a single

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

SpurAudio: A Benchmark for Studying Shortcut Learning in Few-Shot Audio Classification

DGX agent

arXiv:2605.13672v1 Announce Type: new Abstract: Few-shot classification (FSC) is widely used for learning from limited labeled data, yet most evaluations implicitly assume that target concepts are ind

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

State-Space NTK Collapse Near Bifurcations

DGX agent

arXiv:2605.12763v1 Announce Type: new Abstract: Rich feature learning in tasks that unfold over time often requires the model to pass through bifurcations, constituting qualitative changes in the unde

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

Stochastic Dimension-Free Zeroth-Order Estimator for High-Dimensional and High-Order PINNs

DGX agent

arXiv:2603.24002v2 Announce Type: replace Abstract: Physics-Informed Neural Networks (PINNs) for high-dimensional and high-order partial differential equations (PDEs) are primarily constrained by the

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

StreamGaze: Gaze-Guided Temporal Reasoning and Proactive Understanding in Streaming Videos

DGX agent

arXiv:2512.01707v3 Announce Type: replace-cross Abstract: Streaming video understanding requires models not only to process temporally incoming frames, but also to anticipate user intention for realis

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Stress-Testing the Reasoning Competence of LLMs With Proofs Under Minimal Formalism

DGX agent

arXiv:2605.12524v1 Announce Type: cross Abstract: We introduce ProofGrid, a benchmark suite for evaluating LLM reasoning through machine-checkable proofs rather than final answers alone. ProofGrid con

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

SupChain-Bench: Benchmarking Large Language Models for Real-World Supply Chain Management

DGX agent

arXiv:2602.07342v2 Announce Type: replace Abstract: Large language models (LLMs) have shown promise in complex reasoning and tool-based decision making, motivating their application to real-world supp

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

SynCABEL: Synthetic Contextualized Augmentation for Biomedical Entity Linking

DGX agent

arXiv:2601.19667v2 Announce Type: replace-cross Abstract: We present SynCABEL (Synthetic Contextualized Augmentation for Biomedical Entity Linking), a framework that addresses a central bottleneck in

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Systematic Failures in Collective Reasoning under Distributed Information in Multi-Agent LLMs

DGX agent

arXiv:2505.11556v4 Announce Type: replace-cross Abstract: Multi-agent systems built on large language models (LLMs) are expected to enhance decision-making by pooling distributed information, yet syst

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Table-R1: Region-based Reinforcement Learning for Table Understanding

DGX agent

arXiv:2505.12415v3 Announce Type: replace-cross Abstract: Tables present unique challenges for language models due to their structured row-column interactions, necessitating specialized approaches for

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

The critical slowing down in diffusion models

DGX agent

arXiv:2605.12597v1 Announce Type: cross Abstract: Computational sampling has been central to the sciences since the mid-20th century. While machine-learning-based approaches have recently enabled majo

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

The Geometry of LLM Quantization: GPTQ as Babai's Nearest Plane Algorithm

DGX agent

arXiv:2507.18553v4 Announce Type: replace Abstract: Quantizing the weights of large language models (LLMs) from 16-bit to lower bitwidth is the de facto approach to deploy massive transformers onto mo

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

TiCo: Time-Controllable Spoken Dialogue Model

DGX agent

arXiv:2603.22267v2 Announce Type: replace-cross Abstract: We introduce TiCo, a time-controllable spoken dialogue model (SDM) that follows time-constrained instructions (e.g., 'Please generate a respon

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Tighter Learning Guarantees on Digital Computers via Concentration of Measure on Finite Spaces

DGX agent

arXiv:2402.05576v4 Announce Type: replace Abstract: Machine learning models with inputs in a Euclidean space R^d, when implemented on digital computers, generalize, and their generalization gap conver

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

TokaMind for Power Grid: Cross-Domain Transfer from Fusion Plasma

DGX agent

arXiv:2605.11033v1 Announce Type: cross Abstract: TokaMind is a multi-modal transformer (MMT) foundation model pre-trained on tokamak plasma diagnostics data from MAST, where it was shown to outperfor

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

ToolWeave: Structured Synthesis of Complex Multi-Turn Tool-Calling Dialogues

DGX agent

arXiv:2605.12521v1 Announce Type: cross Abstract: Multi-turn tool calling is essential for LLMs to function as autonomous agents, yet synthesizing the training data required for these capabilities rem

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Topo-R1: Detecting Topological Anomalies via Vision-Language Models

DGX agent

arXiv:2603.13054v2 Announce Type: replace Abstract: Topology is critical in tubular structures such as blood vessels, nerve fibers, and road networks, where connectivity and loop structure govern down

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

TouchAnything: A Dataset and Framework for Bimanual Tactile Estimation from Egocentric Video

DGX agent

arXiv:2605.13083v1 Announce Type: new Abstract: Egocentric human video data, which captures rich human-environment interactions and can be collected at scale, has become a key driver of embodied intel

model-releasesarxiv-cs-ro
14 May 2026
Model Releases

Toward AI-Driven Digital Twins for Metropolitan Floods: A Conditional Latent Dynamics Network Surrogate of the Shallow Water Equations

DGX agent

arXiv:2605.13761v1 Announce Type: new Abstract: AI-driven flood digital twins demand fast hydrodynamic surrogates for ensemble forecasting and observation assimilation. Yet even GPU-accelerated two-di

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

Toward Scalable Verifiable Reward: Proxy State-Based Evaluation for Multi-turn Tool-Calling LLM Agents

DGX agent

arXiv:2602.16246v3 Announce Type: replace Abstract: Interactive large language model (LLM) agents operating via multi-turn dialogue and multi-step tool calling are increasingly used in production. Ben

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Towards Long-horizon Embodied Agents with Tool-Aligned Vision-Language-Action Models

DGX agent

arXiv:2605.13119v1 Announce Type: cross Abstract: Vision-language-action (VLA) models are effective robot action executors, but they remain limited on long-horizon tasks due to the dual burden of exte

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Training Large Language Models to Predict Clinical Events

DGX agent

arXiv:2605.12817v1 Announce Type: cross Abstract: Longitudinal clinical notes contain rich evidence of how patients evolve over time, but converting this signal into training supervision for clinical

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Training LLMs with Reinforcement Learning for Intent-Aware Personalized Question Answering

DGX agent

arXiv:2605.12645v1 Announce Type: cross Abstract: Effective personalized question answering (PQA) in language models requires grounding responses in the user's underlying intent, where intent refers t

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Uncertainty-Aware Prediction of Lung Tumor Growth from Sparse Longitudinal CT Data via Bayesian Physics-Informed Neural Networks

DGX agent

arXiv:2605.13560v1 Announce Type: new Abstract: This work studies lung tumor growth prediction from sparse and irregular longitudinal computed tomography (CT) observations with measurement variability

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

Uncertainty-Driven Anomaly Detection for Psychotic Relapse Using Smartwatches: Forecasting and Multi-Task Learning Fusion

DGX agent

arXiv:2605.13816v1 Announce Type: new Abstract: Digital phenotyping enables continuous passive monitoring of behavior and physiology, offering a promising paradigm for early detection of psychotic rel

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

Understanding and Accelerating the Training of Masked Diffusion Language Models

DGX agent

arXiv:2605.13026v1 Announce Type: cross Abstract: Masked diffusion models (MDMs) have emerged as a promising alternative to autoregressive models (ARMs) for language modeling. However, MDMs are known

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Understanding Catastrophic Forgetting In LoRA via Mean-Field Attention Dynamics

DGX agent

arXiv:2402.15415v2 Announce Type: replace Abstract: Low-Rank Adaptation (LoRA) is the dominant parameter-efficient fine-tuning method due to its favorable compute-performance trade-off, yet it suffers

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

UNIV: Unified Foundation Model for Infrared and Visible Modalities

DGX agent

arXiv:2509.15642v3 Announce Type: replace Abstract: Joint RGB-infrared perception is essential for achieving robustness under diverse weather and illumination conditions. Although foundation models ex

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

Universal Representation of Generalized Convex Functions and their Gradients

DGX agent

arXiv:2509.04477v3 Announce Type: replace-cross Abstract: A wide range of optimization problems can often be written in terms of generalized convex functions (GCFs). When this structure is present, it

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

Unlocking Patch-Level Features for CLIP-Based Class-Incremental Learning

DGX agent

arXiv:2605.13835v1 Announce Type: new Abstract: Class-Incremental Learning (CIL) enables models to continuously integrate new knowledge while mitigating catastrophic forgetting. Driven by the remarkab

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

Useful Memories Become Faulty When Continuously Updated by LLMs

DGX agent

arXiv:2605.12978v1 Announce Type: new Abstract: Learning from past experience benefits from two complementary forms of memory: episodic traces -- raw trajectories of what happened -- and consolidated

model-releasesarxiv-cs-ai
14 May 2026
← Previous
1…244245246247248…361
Next →