AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,778 results
Model Releases

Physics Question Scene Graph: Fine-grained Evaluation of Physical Plausibility in Text-to-Video Generation

DGX agent

arXiv:2606.25306v1 Announce Type: new Abstract: Video generation models are increasingly capable of producing realistic videos, but they still struggle to generate videos that follow basic physical la

model-releasesarxiv-cs-cv
25 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models

DGX agent

arXiv:2606.25442v1 Announce Type: new Abstract: Safety alignment of large language models (LLMs) typically depends on high-quality supervision data, such as safe demonstrations or preference pairs. Ho

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

Pre-Warm: Input-Conditioned Weight Initialization for Convolutional Neural Networks

DGX agent

arXiv:2606.25256v1 Announce Type: new Abstract: We introduce Pre-Warm, a simple yet effective zero-training-cost method for data-conditioned initialization of the first convolutional layer. Before the

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

PRISM: Feed-Forward Single-Image 3D Reconstruction via Geometric Warp-Residual Modeling

DGX agent

arXiv:2606.25430v1 Announce Type: new Abstract: Reconstructing 3D scenes from a single image is a fundamental challenge in computer vision, with broad applications in virtual reality, robotics, and co

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

Privacy-Aware Visual Language Models

DGX agent

arXiv:2405.17423v4 Announce Type: replace-cross Abstract: As Visual Language Models (VLMs) become increasingly embedded in everyday applications, ensuring they can recognise and appropriately handle p

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

Project Auto-World: Towards Automated Benchmarking of Neural Relational Reasoners

DGX agent

arXiv:2606.24965v1 Announce Type: cross Abstract: Reasoning about relational structures remains a significant challenge for neural models, particularly when they must systematically apply learned know

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Pulmonary Embolism Risk Stratification from CTPA and Medical Records: Vascular Graphs Are Not All You Need

DGX agent

arXiv:2606.25956v1 Announce Type: new Abstract: Risk stratification for pulmonary embolism (PE) is critical for clinical decision-making. Stratification guidelines are based on patient medical records

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

PVF:Understanding AI Vulnerability Against SDCs

DGX agent

arXiv:2405.01741v4 Announce Type: replace-cross Abstract: Reliability of AI systems is a fundamental concern for the successful deployment and widespread adoption of AI technologies. Unfortunately, th

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

RAS: Measuring LLM Safety Through Refusal Alignment

DGX agent

arXiv:2606.25750v1 Announce Type: cross Abstract: Safety evaluation of large language models (LLMs) is commonly performed by querying models with unsafe or jailbreak prompts and judging whether their

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

Rational Neural Networks have Expressivity Advantages

DGX agent

arXiv:2602.12390v2 Announce Type: replace Abstract: We study neural networks with trainable low-degree rational activation functions and show that they are more expressive and parameter-efficient than

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Real-Time Voice AI Hears but Does Not Listen

DGX agent

arXiv:2606.26083v1 Announce Type: new Abstract: Speech conveys information through both words and vocal delivery. We evaluate four leading production realtime voice systems-OpenAI's GPT Realtime 2, Go

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

Reasonable Motion: A General ASP Foundation for Environment Constrained Movement Trajectory Computation

DGX agent

arXiv:2606.25626v1 Announce Type: cross Abstract: We present a general answer set programming based hybrid quantitative-qualitative method for computing constrained branching trajectory modes for movi

model-releasesarxiv-cs-ro
25 Jun 2026
Model Releases

Reclaim Evaluation: A Lossy Memory Is Worse Than an Empty One

DGX agent

arXiv:2606.25449v1 Announce Type: new Abstract: A language model's memory can be worse than having no memory at all. Give a model a memory that kept a wrong conclusion but dropped the work behind it,

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

RevengeBench: Reverse Engineering Code-Space Policies from Behavioral Experiments

DGX agent

arXiv:2606.26094v1 Announce Type: new Abstract: For most of scientific history, researchers studying behavior could only infer hidden mechanisms from outward actions: an inverse problem that becomes m

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

RigPI: Dynamic Parameter Identification of Rigid Body via VLM-Seeded Differentiable Simulation

DGX agent

arXiv:2606.25212v1 Announce Type: new Abstract: Accurate physical parameter identification of manipulated objects is fundamental to advanced robotic manipulation and the construction of faithful digit

model-releasesarxiv-cs-ro
25 Jun 2026
Model Releases

RL fine-tuning is now live for @nvidiaai Nemotron 3 on Fireworks, starting with Nemotron 3 Super (LoRA). Train with GRPO and serve the model…

DGX agent

RL fine-tuning is now live for @nvidiaai Nemotron 3 on Fireworks, starting with Nemotron 3 Super (LoRA). Train with GRPO and serve the model in one place. We price by GPU-hour, not per token, so long

model-releasesfireworks-ai--x
25 Jun 2026
Model Releases

RoboAtlas: Contextual Active SLAM

DGX agent

arXiv:2606.26046v1 Announce Type: cross Abstract: We present RoboAtlas, a contextual Active SLAM framework that adaptively balances geometric exploration and semantic reasoning using a scalable 3D sem

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

RoboRouter: Training-Free Policy Routing for Robotic Manipulation

DGX agent

arXiv:2603.07892v4 Announce Type: replace Abstract: Research on robotic manipulation has developed a diverse set of policy paradigms, including vision-language-action (VLA) models, vision-action (VA)

model-releasesarxiv-cs-ro
25 Jun 2026
Model Releases

Robustness assessment of large audio language models in multiple-choice evaluation

DGX agent

arXiv:2510.04584v2 Announce Type: replace Abstract: Recent advances in large audio language models (LALMs) have primarily been assessed using a multiple-choice question answering (MCQA) framework. How

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

RWGBench: Evaluating Scholarly Positioning in Related Work Generation

DGX agent

arXiv:2606.24894v1 Announce Type: cross Abstract: Large language models have shown strong fluency in scientific writing, yet the evaluation of related work generation (RWG) remains limited. Existing R

model-releasesarxiv-cs-ai
25 Jun 2026
Model Releases

Salesforce launches Help Agent to simplify AI customer service deployment

DGX agent

Salesforce Inc. is launching a new prepackaged artificial intelligence agent for customer service, enabling organizations to quickly build and deploy AI agents. Today Salesforce announced Help Agent,

model-releasessiliconangle
25 Jun 2026
Model Releases

Same Evidence, Different Answer: Auditing Order Sensitivity in Multimodal Large Language Models

DGX agent

arXiv:2606.26079v1 Announce Type: new Abstract: Standard benchmarks for multimodal large language models (MLLMs) score each item on one canonical ordering and miss whether order-irrelevant shuffling c

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

Sampling Strategies for Robust Universal Quadrupedal Locomotion Policies

DGX agent

arXiv:2510.07094v2 Announce Type: replace Abstract: This work focuses on sampling strategies of configuration variations for generating robust universal locomotion policies for quadrupedal robots. We

model-releasesarxiv-cs-ro
25 Jun 2026
Model Releases

SARA: Unlocking Multilingual Knowledge in Mixture-of-Experts via Semantically Anchored Routing Alignment

DGX agent

arXiv:2606.25821v1 Announce Type: new Abstract: Sparse Mixture-of-Experts (MoE) architectures have emerged as an increasingly influential paradigm as they offer a strategic balance between parameter s

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

Sarashina2.2-TTS: Tackling Kanji Polyphony in Japanese Speech Generation via Data Scaling and Targeted Data Synthesis

DGX agent

arXiv:2606.25369v1 Announce Type: cross Abstract: While large language model (LLM)-based text-to-speech (TTS) systems have achieved high-quality speech synthesis, most existing systems focus on Englis

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

SciRisk-Bench: A Risk-Dimension-Aware Benchmark for AI4Science Safety

DGX agent

arXiv:2606.18936v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly embedded in AI for Science (AI4Science) workflows, from scientific question answering and literature a

model-releasesarxiv-cs-ai
25 Jun 2026
Model Releases

SFL-MTSC: Leveraging Semantic Frame-Level Multi-Task Self-Consistency for Robust Multi-Intent Spoken Language Understanding

DGX agent

arXiv:2606.25552v1 Announce Type: new Abstract: Prompt-based spoken language understanding (SLU) with large language models (LLMs) often suffers from inconsistent intent--slot structures due to decodi

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

Shapley-Inspired Feature Weighting in k-means with No Additional Hyperparameters

DGX agent

arXiv:2508.07952v2 Announce Type: replace Abstract: Clustering algorithms often assume all features contribute equally to the data structure, an assumption that usually fails in high-dimensional or no

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

ShutterMuse: Capture-Time Photography Guidance with MLLMs

DGX agent

arXiv:2606.25763v1 Announce Type: new Abstract: Real-world photography requires capture-time guidance for both camera framing and subject pose. Yet existing aesthetic cropping benchmarks mainly evalua

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

Silent Failures in Physics-Informed Neural Networks: Parameter Poisoning and the Limits of Loss-Based Validation

DGX agent

arXiv:2606.25151v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) embed governing equations in their loss function, enabling mesh-free solutions to partial differential equation

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Small edits, large models: How Wikipedia advocacy shapes LLM values

DGX agent

arXiv:2606.24890v1 Announce Type: new Abstract: Can a small group of volunteers shape how AI systems discuss animal welfare, just by editing Wikipedia? We show that they can. Wikipedia appears in near

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

Small Initialization Matters for Large Language Models

DGX agent

arXiv:2606.17945v2 Announce Type: replace Abstract: Large language models provide a tractable system for asking how intelligence itself emerges, rather than only how LLMs can be engineered. Although p

model-releasesarxiv-cs-ai
25 Jun 2026
Model Releases

SpaceX rocket launches in 2026 are insanely brutal SpaceX: • 76 operational launches ➝ 76 successes ➝ 0 failures The entire rest of the worl…

DGX agent

SpaceX rocket launches in 2026 are insanely brutal SpaceX: • 76 operational launches ➝ 76 successes ➝ 0 failures The entire rest of the world combined: • 55 tracked launches ➝ 49 successes ➝ 6 failure

model-releaseselon-musk--x
25 Jun 2026
Model Releases

SPARC: Separating Perception And Reasoning Circuits for Test-time Scaling of VLMs

DGX agent

arXiv:2602.06566v3 Announce Type: replace-cross Abstract: Despite recent successes, test-time scaling -- i.e., dynamically expanding the token budget during inference as needed -- remains brittle for

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

Spatio-Temporal Mixture-of-Modality-Experts Diffusion for Quantitative DCE-MRI Synthesis from Incomplete MR Sequences

DGX agent

arXiv:2606.25535v1 Announce Type: new Abstract: Quantitative maps from dynamic contrast-enhanced MRI (DCE-MRI) are essential for tumor assessment but are often unavailable due to contrast-agent risks

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

Speech Codec Probing from Semantic and Phonetic Perspectives

DGX agent

arXiv:2603.10371v2 Announce Type: replace-cross Abstract: Speech tokenizers are essential for connecting speech to large language models (LLMs) in multimodal systems. Speech tokenizers are expected to

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

SpeechEQ: Benchmarking Emotional Intelligence Quotient in Socially Aware Voice Conversational Models

DGX agent

arXiv:2606.25990v1 Announce Type: new Abstract: As multimodal conversational systems increasingly engage in spoken interaction, their ability to navigate paralinguistic social cues has become a critic

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

Spotlighting Task-Relevant Features: Object-Centric Representations for Better Generalization in Robotic Manipulation

DGX agent

arXiv:2601.21416v2 Announce Type: replace Abstract: The generalization capabilities of robotic manipulation policies are heavily influenced by the choice of visual representations. Existing approaches

model-releasesarxiv-cs-ro
25 Jun 2026
Model Releases

SSMNBench: Diagnosing Image-based Cross-View Human-Object Understanding via Single-View Sufficiency and Multi-View Necessity

DGX agent

arXiv:2606.25634v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have shown remarkable progress in single-image perception, yet their ability to reason about complex cross-view

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

Stable-Shift: Biologically Structured Prediction of Transcriptional Responses to Unseen Gene Perturbations

DGX agent

arXiv:2606.24940v1 Announce Type: cross Abstract: Predicting transcriptional responses to genetic perturbations could reduce the experimental burden of functional genomics, but extrapolation to genes

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Stage-Aware and Roughness-Constrained Diffusion Policy for Multi-Stage Robotic Polishing

DGX agent

arXiv:2606.25754v1 Announce Type: new Abstract: Polishing is a critical finishing process in high-end manufacturing fields such as aerospace, where surface quality directly affects the service perform

model-releasesarxiv-cs-ro
25 Jun 2026
Model Releases

Staying In Character: Perspective-Bounded Memory For Book-Based Role-Playing Agents

DGX agent

arXiv:2606.25632v1 Announce Type: new Abstract: Recent LLM role-playing systems build character agents from novels by extracting characters, scenes, and relations. Yet long-narrative role-playing suff

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

STEB: A Speech-to-Speech Translation Expressiveness Benchmark for Evaluating Beyond Translation Fidelity

DGX agent

arXiv:2606.25529v1 Announce Type: cross Abstract: Speech-to-speech translation (S2ST) should preserve not only lexical meaning, but also expressive attributes: emotion, scenario style (e.g., news repo

model-releasesarxiv-cs-ai
25 Jun 2026
Model Releases

Steering Vision-Language Models with Joint Sparse Autoencoders

DGX agent

arXiv:2606.25657v1 Announce Type: new Abstract: Sparse Autoencoders (SAEs) have shown promise for analyzing language models, but applying them to vision-language models (VLMs) often yields representat

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

Story Operators: Decomposing the Original o Sequel Transformation in Embedding Space

DGX agent

arXiv:2606.25379v1 Announce Type: new Abstract: I treat a book as a point in a sentence-embedding space and a literary transformation as an operation on points. Given an original novel and its sequel,

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

SurgAtlas: A Large-Scale Surgical Video-Language Dataset with 2,391 Hours of Open and Minimally Invasive Surgery

DGX agent

arXiv:2606.25905v1 Announce Type: new Abstract: We introduce SurgAtlas, the largest surgical video-language dataset to date, comprising 15,291 videos (2,391 hours) spanning 18 surgical specialties and

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

Swazure: Swarm Measurement of Pose for Flying Light Specks

DGX agent

arXiv:2606.25222v1 Announce Type: new Abstract: One may construct a 3D multimedia display using miniature drones configured with light sources, Flying Light Specks (FLSs). Swarms of FLSs localize to i

model-releasesarxiv-cs-ro
25 Jun 2026
Model Releases

SyncLoop: A Multimodal Dual-Loop Framework for Self-Improving Mathematical Reasoning

DGX agent

arXiv:2507.16518v3 Announce Type: replace-cross Abstract: Recent advances in multimodal large language models (MLLMs) have shown impressive reasoning capabilities. However, further enhancing existing

model-releasesarxiv-cs-cl
25 Jun 2026
← Previous
1…171172173174175…475
Next →