AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,585 results
25 Jun 2026

LLMs are getting better at writing GPU kernels. Multi-GPU kernels are the harder test. At @aiDotEngineer World's Fair, @simran_s_arora will …

Model ReleasesDGX agent

LLMs are getting better at writing GPU kernels. Multi-GPU kernels are the harder test. At @aiDotEngineer World's Fair, @simran_s_arora will share ParallelKernelBench, an open-source benchmark built fr

MacroLens: A Multi-Task Benchmark for Contextual Financial Reasoning under Macroeconomic Scenarios

Model ReleasesDGX agent

arXiv:2606.24950v1 Announce Type: new Abstract: Financial decision-making is contextual: forecasting prices, valuing companies, and assessing event exposure weigh price history, accounting fundamental

Make Watergate Great Again.

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Make Watergate Great Again. JD Vance: 'I think Nixon's historical legacy is enjoying a bit of a renaissance, and deservedly so. I joked that if Watergate happened tomorrow, it would be like a 12 hours

MedLayBench-V: A Large-Scale Benchmark for Expert-Lay Semantic Alignment in Medical Vision Language Models

Model ReleasesDGX agent

arXiv:2604.05738v2 Announce Type: replace Abstract: Medical Vision-Language Models (Med-VLMs) have achieved expert-level proficiency in interpreting diagnostic imaging. However, current models are pre

Memory-Efficient Policy Libraries with Low-Rank Adaptation in Reinforcement Learning

Model ReleasesDGX agent

arXiv:2606.25700v1 Announce Type: new Abstract: When fine-tuning Large Language Models (LLMs), there has been success in minimizing both memory usage and computation with Parameter-Efficient Fine-Tuni

MINIF2F-DAFNY: LLM-Guided Mathematical Theorem Proving via Auto-Active Verification

Model ReleasesDGX agent

arXiv:2512.10187v3 Announce Type: replace Abstract: LLMs excel at reasoning, but validating their steps remains challenging. Formal verification offers a solution through mechanically checkable proofs

Model-agnostic Mitigation Strategies of Data Imbalance for Regression

Model ReleasesDGX agent

arXiv:2506.01486v2 Announce Type: replace Abstract: Data imbalance persists as a pervasive challenge in regression tasks, introducing bias in model performance and undermining predictive reliability.

Model Forensics: Investigating Whether Concerning Behavior Reflects Misalignment

Model ReleasesDGX agent

arXiv:2606.26071v1 Announce Type: new Abstract: A central goal of safety research is determining whether a model is misaligned. Prior work has largely focused on detecting concerning behavior. But beh

Multi-Agent Goal Recognition with Team- and Goal-Conditioned Reinforcement Learning and Factorized Branch-and-Bound

Model ReleasesDGX agent

arXiv:2606.25978v1 Announce Type: cross Abstract: Multi-agent goal recognition asks an observer to jointly infer which agents act together and what each team is trying to achieve, so the hypothesis sp

Multi-Stream Temporal Fusion for Financial Fraud Detection

Model ReleasesDGX agent

arXiv:2606.25007v1 Announce Type: new Abstract: Financial fraud detection in digital banking requires reasoning over multiple heterogeneous event streams -- transactions, login sessions, risk signals

Multilingual Hematology Visual Question Answering Dataset

Model ReleasesDGX agent

arXiv:2606.25246v1 Announce Type: cross Abstract: Vision Language Models (VLMs) have shown promising capabilities in medical image analysis by jointly understanding visual and textual information for

Natural Ungrokking: Asymmetric Control of Which Rules Survive Pretraining

Model ReleasesDGX agent

arXiv:2606.26050v1 Announce Type: cross Abstract: Midway through an ordinary pretraining run, a small language model learns the pronoun-gender rule: cued with a girl's name ('Sue cried because'), it r

Ok, I am interviewing the legend @trq212 on Friday. I'm planning to ask him to show me: → His Claude Code setup and how he uses /goal and /l…

Model ReleasesDGX agent

Ok, I am interviewing the legend @trq212 on Friday. I'm planning to ask him to show me: → His Claude Code setup and how he uses /goal and /loop and dynamic workflows → How he does planning with HTML +

Omni-Perception Policy Optimization for Multimodal Emotion Reasoning

Model ReleasesDGX agent

arXiv:2606.25325v1 Announce Type: new Abstract: We find that current emotion-oriented Omni-MLLMs still lack reliable omni-modal perception: they (i) underutilize multimodal cues in their reasoning tra

On-Device Neural Architecture Search

Model ReleasesDGX agent

arXiv:2606.24900v1 Announce Type: new Abstract: This paper proposes a new approach to near-sensor computing, in which a lightweight Neural Architecture Search (NAS) is performed directly on the deploy

OpenAI staggers GPT-5.6 rollout for government vetting, eyes 2027 IPO

Model ReleasesDGX agent

OpenAI Group PBC will roll out its next model, GPT-5.6, to a small group of partners rather than the public at the request of the Trump administration, the latest sign that Washington now wants to rev

OpenAI will delay GPT-5.6 after Trump administration request

Model ReleasesDGX agent

The Trump administration, apprehensive of potential security issues, has reportedly asked OpenAI to stagger the release of its next big-ticket model, GPT-5.6. The Information reported that OpenAI CEO

Operator Boosting Produces Pareto-Efficient PDE Surrogates

Model ReleasesDGX agent

arXiv:2606.17460v2 Announce Type: replace Abstract: Neural operators are widely used as surrogate solution maps for partial differential equations (PDEs), but full-size models can be costly to store,

OracleAnalyser: Analysing Implicit Semantics of Oracle Bone Scripts through MLLMs with Post-training

Model ReleasesDGX agent

arXiv:2606.25906v1 Announce Type: new Abstract: With the advancement of artificial intelligence, research on oracle bone scripts has entered a new era. However, existing methods and benchmarks remain

OrthoTrack: Continuous 6-DoF UAV Trajectory Estimation Anchored in Public Orthophotos

Model ReleasesDGX agent

arXiv:2606.25245v1 Announce Type: new Abstract: Continuous 6-DoF pose estimation is essential for autonomous UAV operations. Yet, existing visual odometry and SLAM methods accumulate drift and yield o

Our new AI policy is that the White House decides ad hoc, for whatever reasons it likes, who does and does not get access to frontier intell…

Model ReleasesDGX agent

Our new AI policy is that the White House decides ad hoc, for whatever reasons it likes, who does and does not get access to frontier intelligence. This seems rather maximally terrible. The US Governm

PatchINR: Patch-Based Implicit Neural Representations for Efficient and Scalable Inference

Model ReleasesDGX agent

arXiv:2606.25534v1 Announce Type: new Abstract: Implicit Neural Representation (INR) provides an effective approach for continuous signal modeling, but classical per-pixel inference results in quadrat

Perfect Detection, Failed Control: The Geometry of Knowing vs. Steering in Language Models

Model ReleasesDGX agent

arXiv:2606.24952v1 Announce Type: new Abstract: A central aspiration of mechanistic interpretability is controllability: if we know where a behavior is represented in a model's activations, we should

Physics Question Scene Graph: Fine-grained Evaluation of Physical Plausibility in Text-to-Video Generation

Model ReleasesDGX agent

arXiv:2606.25306v1 Announce Type: new Abstract: Video generation models are increasingly capable of producing realistic videos, but they still struggle to generate videos that follow basic physical la

PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models

Model ReleasesDGX agent

arXiv:2606.25442v1 Announce Type: new Abstract: Safety alignment of large language models (LLMs) typically depends on high-quality supervision data, such as safe demonstrations or preference pairs. Ho

Pre-Warm: Input-Conditioned Weight Initialization for Convolutional Neural Networks

Model ReleasesDGX agent

arXiv:2606.25256v1 Announce Type: new Abstract: We introduce Pre-Warm, a simple yet effective zero-training-cost method for data-conditioned initialization of the first convolutional layer. Before the

PRISM: Feed-Forward Single-Image 3D Reconstruction via Geometric Warp-Residual Modeling

Model ReleasesDGX agent

arXiv:2606.25430v1 Announce Type: new Abstract: Reconstructing 3D scenes from a single image is a fundamental challenge in computer vision, with broad applications in virtual reality, robotics, and co

Privacy-Aware Visual Language Models

Model ReleasesDGX agent

arXiv:2405.17423v4 Announce Type: replace-cross Abstract: As Visual Language Models (VLMs) become increasingly embedded in everyday applications, ensuring they can recognise and appropriately handle p

Project Auto-World: Towards Automated Benchmarking of Neural Relational Reasoners

Model ReleasesDGX agent

arXiv:2606.24965v1 Announce Type: cross Abstract: Reasoning about relational structures remains a significant challenge for neural models, particularly when they must systematically apply learned know

Pulmonary Embolism Risk Stratification from CTPA and Medical Records: Vascular Graphs Are Not All You Need

Model ReleasesDGX agent

arXiv:2606.25956v1 Announce Type: new Abstract: Risk stratification for pulmonary embolism (PE) is critical for clinical decision-making. Stratification guidelines are based on patient medical records

PVF:Understanding AI Vulnerability Against SDCs

Model ReleasesDGX agent

arXiv:2405.01741v4 Announce Type: replace-cross Abstract: Reliability of AI systems is a fundamental concern for the successful deployment and widespread adoption of AI technologies. Unfortunately, th

RAS: Measuring LLM Safety Through Refusal Alignment

Model ReleasesDGX agent

arXiv:2606.25750v1 Announce Type: cross Abstract: Safety evaluation of large language models (LLMs) is commonly performed by querying models with unsafe or jailbreak prompts and judging whether their

Rational Neural Networks have Expressivity Advantages

Model ReleasesDGX agent

arXiv:2602.12390v2 Announce Type: replace Abstract: We study neural networks with trainable low-degree rational activation functions and show that they are more expressive and parameter-efficient than

Real-Time Voice AI Hears but Does Not Listen

Model ReleasesDGX agent

arXiv:2606.26083v1 Announce Type: new Abstract: Speech conveys information through both words and vocal delivery. We evaluate four leading production realtime voice systems-OpenAI's GPT Realtime 2, Go

Reasonable Motion: A General ASP Foundation for Environment Constrained Movement Trajectory Computation

Model ReleasesDGX agent

arXiv:2606.25626v1 Announce Type: cross Abstract: We present a general answer set programming based hybrid quantitative-qualitative method for computing constrained branching trajectory modes for movi

Reclaim Evaluation: A Lossy Memory Is Worse Than an Empty One

Model ReleasesDGX agent

arXiv:2606.25449v1 Announce Type: new Abstract: A language model's memory can be worse than having no memory at all. Give a model a memory that kept a wrong conclusion but dropped the work behind it,

RevengeBench: Reverse Engineering Code-Space Policies from Behavioral Experiments

Model ReleasesDGX agent

arXiv:2606.26094v1 Announce Type: new Abstract: For most of scientific history, researchers studying behavior could only infer hidden mechanisms from outward actions: an inverse problem that becomes m

RigPI: Dynamic Parameter Identification of Rigid Body via VLM-Seeded Differentiable Simulation

Model ReleasesDGX agent

arXiv:2606.25212v1 Announce Type: new Abstract: Accurate physical parameter identification of manipulated objects is fundamental to advanced robotic manipulation and the construction of faithful digit

RL fine-tuning is now live for @nvidiaai Nemotron 3 on Fireworks, starting with Nemotron 3 Super (LoRA). Train with GRPO and serve the model…

Model ReleasesDGX agent

RL fine-tuning is now live for @nvidiaai Nemotron 3 on Fireworks, starting with Nemotron 3 Super (LoRA). Train with GRPO and serve the model in one place. We price by GPU-hour, not per token, so long

RoboAtlas: Contextual Active SLAM

Model ReleasesDGX agent

arXiv:2606.26046v1 Announce Type: cross Abstract: We present RoboAtlas, a contextual Active SLAM framework that adaptively balances geometric exploration and semantic reasoning using a scalable 3D sem

RoboRouter: Training-Free Policy Routing for Robotic Manipulation

Model ReleasesDGX agent

arXiv:2603.07892v4 Announce Type: replace Abstract: Research on robotic manipulation has developed a diverse set of policy paradigms, including vision-language-action (VLA) models, vision-action (VA)

Robustness assessment of large audio language models in multiple-choice evaluation

Model ReleasesDGX agent

arXiv:2510.04584v2 Announce Type: replace Abstract: Recent advances in large audio language models (LALMs) have primarily been assessed using a multiple-choice question answering (MCQA) framework. How

RWGBench: Evaluating Scholarly Positioning in Related Work Generation

Model ReleasesDGX agent

arXiv:2606.24894v1 Announce Type: cross Abstract: Large language models have shown strong fluency in scientific writing, yet the evaluation of related work generation (RWG) remains limited. Existing R

Salesforce launches Help Agent to simplify AI customer service deployment

Model ReleasesDGX agent

Salesforce Inc. is launching a new prepackaged artificial intelligence agent for customer service, enabling organizations to quickly build and deploy AI agents. Today Salesforce announced Help Agent,

Same Evidence, Different Answer: Auditing Order Sensitivity in Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2606.26079v1 Announce Type: new Abstract: Standard benchmarks for multimodal large language models (MLLMs) score each item on one canonical ordering and miss whether order-irrelevant shuffling c

Sampling Strategies for Robust Universal Quadrupedal Locomotion Policies

Model ReleasesDGX agent

arXiv:2510.07094v2 Announce Type: replace Abstract: This work focuses on sampling strategies of configuration variations for generating robust universal locomotion policies for quadrupedal robots. We

SARA: Unlocking Multilingual Knowledge in Mixture-of-Experts via Semantically Anchored Routing Alignment

Model ReleasesDGX agent

arXiv:2606.25821v1 Announce Type: new Abstract: Sparse Mixture-of-Experts (MoE) architectures have emerged as an increasingly influential paradigm as they offer a strategic balance between parameter s

Sarashina2.2-TTS: Tackling Kanji Polyphony in Japanese Speech Generation via Data Scaling and Targeted Data Synthesis

Model ReleasesDGX agent

arXiv:2606.25369v1 Announce Type: cross Abstract: While large language model (LLM)-based text-to-speech (TTS) systems have achieved high-quality speech synthesis, most existing systems focus on Englis

SciRisk-Bench: A Risk-Dimension-Aware Benchmark for AI4Science Safety

Model ReleasesDGX agent

arXiv:2606.18936v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly embedded in AI for Science (AI4Science) workflows, from scientific question answering and literature a

SFL-MTSC: Leveraging Semantic Frame-Level Multi-Task Self-Consistency for Robust Multi-Intent Spoken Language Understanding

Model ReleasesDGX agent

arXiv:2606.25552v1 Announce Type: new Abstract: Prompt-based spoken language understanding (SLU) with large language models (LLMs) often suffers from inconsistent intent--slot structures due to decodi

Shapley-Inspired Feature Weighting in k-means with No Additional Hyperparameters

Model ReleasesDGX agent

arXiv:2508.07952v2 Announce Type: replace Abstract: Clustering algorithms often assume all features contribute equally to the data structure, an assumption that usually fails in high-dimensional or no

ShutterMuse: Capture-Time Photography Guidance with MLLMs

Model ReleasesDGX agent

arXiv:2606.25763v1 Announce Type: new Abstract: Real-world photography requires capture-time guidance for both camera framing and subject pose. Yet existing aesthetic cropping benchmarks mainly evalua

Silent Failures in Physics-Informed Neural Networks: Parameter Poisoning and the Limits of Loss-Based Validation

Model ReleasesDGX agent

arXiv:2606.25151v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) embed governing equations in their loss function, enabling mesh-free solutions to partial differential equation

Small edits, large models: How Wikipedia advocacy shapes LLM values

Model ReleasesDGX agent

arXiv:2606.24890v1 Announce Type: new Abstract: Can a small group of volunteers shape how AI systems discuss animal welfare, just by editing Wikipedia? We show that they can. Wikipedia appears in near

Small Initialization Matters for Large Language Models

Model ReleasesDGX agent

arXiv:2606.17945v2 Announce Type: replace Abstract: Large language models provide a tractable system for asking how intelligence itself emerges, rather than only how LLMs can be engineered. Although p

SpaceX rocket launches in 2026 are insanely brutal SpaceX: • 76 operational launches ➝ 76 successes ➝ 0 failures The entire rest of the worl…

Model ReleasesDGX agent

SpaceX rocket launches in 2026 are insanely brutal SpaceX: • 76 operational launches ➝ 76 successes ➝ 0 failures The entire rest of the world combined: • 55 tracked launches ➝ 49 successes ➝ 6 failure

SPARC: Separating Perception And Reasoning Circuits for Test-time Scaling of VLMs

Model ReleasesDGX agent

arXiv:2602.06566v3 Announce Type: replace-cross Abstract: Despite recent successes, test-time scaling -- i.e., dynamically expanding the token budget during inference as needed -- remains brittle for

Spatio-Temporal Mixture-of-Modality-Experts Diffusion for Quantitative DCE-MRI Synthesis from Incomplete MR Sequences

Model ReleasesDGX agent

arXiv:2606.25535v1 Announce Type: new Abstract: Quantitative maps from dynamic contrast-enhanced MRI (DCE-MRI) are essential for tumor assessment but are often unavailable due to contrast-agent risks

Speech Codec Probing from Semantic and Phonetic Perspectives

Model ReleasesDGX agent

arXiv:2603.10371v2 Announce Type: replace-cross Abstract: Speech tokenizers are essential for connecting speech to large language models (LLMs) in multimodal systems. Speech tokenizers are expected to

SpeechEQ: Benchmarking Emotional Intelligence Quotient in Socially Aware Voice Conversational Models

Model ReleasesDGX agent

arXiv:2606.25990v1 Announce Type: new Abstract: As multimodal conversational systems increasingly engage in spoken interaction, their ability to navigate paralinguistic social cues has become a critic

← Previous
1…133134135136137…377
Next →