AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,574 results
13 May 2026

Efficient and Adaptive Human Activity Recognition via LLM Backbones

Model ReleasesDGX agent

arXiv:2605.12019v1 Announce Type: new Abstract: Human Activity Recognition (HAR) is a core task in pervasive computing systems, where models must operate under strict computational constraints while r

Efficient LLM Reasoning via Variational Posterior Guidance with Efficiency Awareness

Model ReleasesDGX agent

arXiv:2605.11019v1 Announce Type: new Abstract: Although large language models rely on chain-of-thought for complex reasoning, the overthinking phenomenon severely degrades inference efficiency. Exist

EgoEV-HandPose: Egocentric 3D Hand Pose Estimation and Gesture Recognition with Stereo Event Cameras

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.12297v1 Announce Type: new Abstract: Egocentric 3D hand pose estimation and gesture recognition are essential for immersive augmented/virtual reality, human-computer interaction, and roboti

EsoLang-Bench: Evaluating Genuine Reasoning in Large Language Models via Esoteric Programming Languages

Model ReleasesDGX agent

arXiv:2603.09678v2 Announce Type: replace-cross Abstract: Large language models achieve near-ceiling performance on code generation benchmarks, yet most of the programming languages used by popular be

Evaluating the Pre-Consultation Ability of LLMs using Diagnostic Guidelines

Model ReleasesDGX agent

arXiv:2601.03627v3 Announce Type: replace Abstract: We introduce EPAG, a benchmark dataset and framework designed for Evaluating the Pre-consultation Ability of LLMs using diagnostic Guidelines. LLMs

Exact Stiefel Optimization for Probabilistic PLS: Closed-Form Updates, Error Bounds, and Calibrated Uncertainty

Model ReleasesDGX agent

arXiv:2605.11607v1 Announce Type: cross Abstract: Probabilistic partial least squares (PPLS) is a central likelihood-based model for two-view learning when one needs both interpretable latent factors

ExploitGym: Can AI Agents Turn Security Vulnerabilities into Real Attacks?

Model ReleasesDGX agent

arXiv:2605.11086v1 Announce Type: cross Abstract: AI agents are rapidly gaining capabilities that could significantly reshape cybersecurity, making rigorous evaluation urgent. A critical capability is

Extending Kernel Trick to Influence Functions

Model ReleasesDGX agent

arXiv:2605.11239v1 Announce Type: new Abstract: In this paper, we present a dual representation of the influence functions, whose computational complexity scales with dataset size rather than model si

FastUMAP: Scalable Dimensionality Reduction via Bipartite Landmark Sampling

Model ReleasesDGX agent

arXiv:2605.11428v1 Announce Type: new Abstract: Exploratory analysis of high-dimensional data rarely stops at a single embedding. In practice, analysts rerun dimensionality reduction after changing pr

fg-expo: Frontier-guided exploration-prioritized policy optimization via adaptive kl and gaussian curriculum

Model ReleasesDGX agent

arXiv:2605.11403v1 Announce Type: cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has become the standard paradigm for LLM mathematical reasoning, with Group Relative Policy Opti

Finally we're going to get shots of Starship from a deployed payload:

Model ReleasesDGX agent

Finally we're going to get shots of Starship from a deployed payload: Starship’s twelfth flight test will debut the next generation Starship and Super Heavy vehicles, powered by the next evolution of

Fine-Tuning Large Language Models for Cooperative Tactical Deconfliction of Small Unmanned Aerial Systems

Model ReleasesDGX agent

arXiv:2603.28561v2 Announce Type: replace Abstract: The growing deployment of small Unmanned Aerial Systems (sUASs) in low-altitude airspaces has increased the need for reliable tactical deconfliction

Finite Volume-Informed Neural Network Framework for 2D Shallow Water Equations: Rugged Loss Landscapes and the Importance of Data Guidance

Model ReleasesDGX agent

arXiv:2605.11001v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) are a simple surrogate-modelling paradigm for partial differential equations, but their standard strong-form re

FLAME: A New Dataset on FLemish Accounts of Momentary Experiences

Model ReleasesDGX agent

arXiv:2504.14707v3 Announce Type: replace Abstract: We introduce FLAME (FLemish Accounts of Momentary Experiences), a new corpus of nearly 25,000 daily personal narratives in Belgian-Dutch (Flemish),

FLARE: Adaptive Multi-Dimensional Reputation for Robust Client Reliability in Federated Learning

Model ReleasesDGX agent

arXiv:2511.14715v3 Announce Type: replace Abstract: Federated learning (FL) enables collaborative model training while preserving data privacy. However, it remains vulnerable to malicious clients who

Focusable Monocular Depth Estimation

Model ReleasesDGX agent

arXiv:2605.11756v1 Announce Type: new Abstract: Monocular depth foundation models generalize well across scenes, yet they are typically optimized with uniform pixel-wise objectives that do not disting

Freeze Deep, Train Shallow: Interpretable Layer Allocation for Continued Pre-Training

Model ReleasesDGX agent

arXiv:2605.11416v1 Announce Type: new Abstract: Selective layer-wise updates are essential for low-cost continued pre-training of Large Language Models (LLMs), yet determining which layers to freeze o

From Web to Pixels: Bringing Agentic Search into Visual Perception

Model ReleasesDGX agent

arXiv:2605.12497v1 Announce Type: new Abstract: Visual perception connects high-level semantic understanding to pixel-level perception, but most existing settings assume that the decisive evidence for

Generative Diffusion Prior Distillation for Long-Context Knowledge Transfer

Model ReleasesDGX agent

arXiv:2605.11414v1 Announce Type: new Abstract: While traditional time-series classifiers assume full sequences at inference, practical constraints (latency and cost) often limit inputs to partial pre

GeneZip: Region-Aware Compression for Long Context DNA Modeling

Model ReleasesDGX agent

arXiv:2602.17739v3 Announce Type: replace-cross Abstract: Long-context DNA models are limited by token-mixing cost and by how compression allocates representational budget across the genome. Existing

Geometric Factual Recall in Transformers

Model ReleasesDGX agent

arXiv:2605.12426v1 Announce Type: new Abstract: How do transformer language models memorize factual associations? A common view casts internal weight matrices as associative memories over pairs of emb

GeomHerd: A Forward-looking Herding Quantification via Ricci Flow Geometry on Agent Interactive Simulations

Model ReleasesDGX agent

arXiv:2605.11645v1 Announce Type: cross Abstract: Herding -- where agents align their behaviors and act collectively -- is a central driver of market fragility and systemic risk. Existing approaches t

GeoR-Bench: Evaluating Geoscience Visual Reasoning

Model ReleasesDGX agent

arXiv:2605.11541v1 Announce Type: new Abstract: Geoscience intelligence is expected to understand, reason about, and predict earth system changes to support human decision-making in critical domains s

GKnow: Measuring the Entanglement of Gender Bias and Factual Gender

Model ReleasesDGX agent

arXiv:2605.12299v1 Announce Type: new Abstract: Recent works have analyzed the impact of individual components of neural networks on gendered predictions, often with a focus on mitigating gender bias.

Google Named a Leader in the Gartner® Magic Quadrant™ for AI Application Development Platforms: Mid-cycle update

Model ReleasesDGX agent

May 2026 update: We’ve refreshed this post to reflect our mid-cycle positioning and the evolution of our platform since the report was first published last November. Last fall, Google was recognized a

Gradient-Boosted Decision Tree for Listwise Context Model in Multimodal Review Helpfulness Prediction

Model ReleasesDGX agent

arXiv:2305.12678v3 Announce Type: replace Abstract: Multimodal Review Helpfulness Prediction (MRHP) aims to rank product reviews based on predicted helpfulness scores and has been widely applied in e-

GRAFT-ATHENA: Self-Improving Agentic Teams for Autonomous Discovery and Evolutionary Numerical Algorithms

Model ReleasesDGX agent

arXiv:2605.11117v1 Announce Type: new Abstract: Scientific discovery can be modeled as a sequence of probabilistic decisions that map physical problems to numerical solutions. Recent agentic AI system

Grid Games: The Power of Multiple Grids for Quantizing Large Language Models

Model ReleasesDGX agent

arXiv:2605.12327v1 Announce Type: new Abstract: A major recent advance in quantization is given by microscaled 4-bit formats such as NVFP4 and MXFP4, quantizing values into small groups sharing a scal

Grokking or Glitching? How Low-Precision Drives Slingshot Loss Spikes

Model ReleasesDGX agent

arXiv:2605.06152v2 Announce Type: replace-cross Abstract: Deep neural networks exhibit periodic loss spikes during unregularized long-term training, a phenomenon known as the 'Slingshot Mechanism.' Ex

GRP: Goal-Reversed Prompting for Zero-Shot Evaluation with LLMs

Model ReleasesDGX agent

arXiv:2503.06139v2 Announce Type: replace Abstract: Pairwise LLM-as-a-judge evaluation asks the judge to identify the better of two candidate answers. We study a one-line modification that asks for th

gym-invmgmt: An Open Benchmarking Framework for Inventory Management Methods

Model ReleasesDGX agent

arXiv:2605.11355v1 Announce Type: new Abstract: Inventory-policy comparisons are often difficult to interpret because performance depends on the evaluation contract as much as on the policy itself. Di

HE-SNR: Uncovering Latent Logic via Entropy for Guiding Mid-Training on SWE-bench

Model ReleasesDGX agent

arXiv:2601.20255v2 Announce Type: replace-cross Abstract: SWE-bench has emerged as the premier benchmark for evaluating Large Language Models on complex software engineering tasks. While these capabil

HEBATRON: A Hebrew-Specialized Open-Weight Mixture-of-Experts Language Model

Model ReleasesDGX agent

arXiv:2605.11255v1 Announce Type: new Abstract: We present Hebatron, a Hebrew-specialized open-weight large language model built on the NVIDIA Nemotron-3 sparse Mixture-of-Experts architecture. Traini

hey surprise - you can just launch interactive in tmux and then tail the jsonl - shipped a small wrapper...ralph loop iterating to full pari…

Model ReleasesDGX agent

hey surprise - you can just launch interactive in tmux and then tail the jsonl - shipped a small wrapper...ralph loop iterating to full parity rn https://github.com/dexhorthy/shannon Starting June 15,

Hi-GaTA: Hierarchical Gated Temporal Aggregation Adapter for Surgical Video Report Generation

Model ReleasesDGX agent

arXiv:2605.11208v1 Announce Type: new Abstract: Automated, clinician-grade assessment reports for surgical procedures could reduce documentation burden and provide objective feedback, yet remain chall

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer

Model ReleasesDGX agent

arXiv:2605.11061v1 Announce Type: new Abstract: The evolution of visual generative models has long been constrained by fragmented architectures relying on disjoint text encoders and external VAEs. In

Holder Policy Optimisation

Model ReleasesDGX agent

arXiv:2605.12058v1 Announce Type: new Abstract: Group Relative Policy Optimisation (GRPO) enhances large language models by estimating advantages across a group of sampled trajectories. However, mappi

How do you keep Claude working until the job is done? Claude Code helps with this in a few ways, including one we shipped recently: /goal.

Model ReleasesDGX agent

Claude Code includes a `/goal` feature that helps maintain task continuity and ensure Claude continues working until a job is completed. This feature, recently shipped, is one of several ways Claude C

How Glance turns hours of video into mobile-ready clips with AI

Model ReleasesDGX agent

Every day, thousands of hours of new video content sits waiting to be discovered. Most of it lives in long-form, horizontal formats, while audiences are scrolling through vertical feeds on their phone

HTML Artifacts are a big part of how I work with agents now. Artifacts can be more than just static files. When combined with agents, they c…

Model ReleasesDGX agent

HTML Artifacts are a big part of how I work with agents now. Artifacts can be more than just static files. When combined with agents, they can take action or help you take action. This unlocks all kin

Human-Grounded Multimodal Benchmark with 900K-Scale Aggregated Student Response Distributions from Japan's National Assessment of Academic Ability

Model ReleasesDGX agent

arXiv:2605.11663v1 Announce Type: new Abstract: Authentic school examinations provide a high-validity test bed for evaluating multimodal large language models (MLLMs), yet benchmarks grounded in Japan

Ice Cream Doesn't Cause Drowning: Benchmarking LLMs Against Statistical Pitfalls in Causal Inference

Model ReleasesDGX agent

arXiv:2505.13770v3 Announce Type: replace-cross Abstract: Reliable causal inference is essential for making decisions in high-stakes areas like medicine, economics, and public policy. However, it rema

If you use any of the following with your Claude sub, your usage must got cut by 25x: - T3 Code - Conductor - zed - jean - “Claude -p” in yo…

Model ReleasesDGX agent

If you use any of the following with your Claude sub, your usage must got cut by 25x: - T3 Code - Conductor - zed - jean - “Claude -p” in your ci - scripts to call Claude code from other tools They’re

Improving the Accuracy of Amortized Model Comparison with Self-Consistency

Model ReleasesDGX agent

arXiv:2508.20614v3 Announce Type: replace-cross Abstract: Amortized Bayesian model comparison (BMC) enables fast probabilistic ranking of models via simulation-based training of neural surrogates. How

Instagram rolls out Instants, which lets users share ephemeral photos, as an in-app feature in Instagram and as a standalone app in select countries (Zac Hall/9to5Mac)

Model ReleasesDGX agent

Zac Hall / 9to5Mac: Instagram rolls out Instants, which lets users share ephemeral photos, as an in-app feature in Instagram and as a standalone app in select countries — Meta just launched a brand ne

Intention-Conditioned Flow Occupancy Models

Model ReleasesDGX agent

arXiv:2506.08902v4 Announce Type: replace Abstract: Large-scale pre-training has fundamentally changed how machine learning research is done today: large foundation models are trained once, and then c

Investigating simple target-covariate relationships for Chronos-2 and TabPFN-TS

Model ReleasesDGX agent

arXiv:2605.12200v1 Announce Type: new Abstract: Time Series Foundation Models (TSFMs) have recently achieved state-of-the-art performance, often outperforming supervised models in zero-shot settings.

KAN-CL: Per-Knot Importance Regularization for Continual Learning with Kolmogorov-Arnold Networks

Model ReleasesDGX agent

arXiv:2605.12306v1 Announce Type: cross Abstract: Catastrophic forgetting remains the central obstacle in continual learning (CL): parameters shared across tasks interfere with one another, and existi

Keeping Score: Efficiency Improvements in Neural Likelihood Surrogate Training via Score-Augmented Loss Functions

Model ReleasesDGX agent

arXiv:2605.12118v1 Announce Type: cross Abstract: For stochastic process models, parameter inference is often severely bottlenecked by computationally expensive likelihood functions. Simulation-based

KV-Fold: One-Step KV-Cache Recurrence for Long-Context Inference

Model ReleasesDGX agent

arXiv:2605.12471v1 Announce Type: cross Abstract: We introduce KV-Fold, a simple, training-free long-context inference protocol that treats the key-value (KV) cache as the accumulator in a left fold o

Large-Small Model Collaboration for Farmland Semantic Change Detection

Model ReleasesDGX agent

arXiv:2605.12282v1 Announce Type: new Abstract: Farmland Semantic Change Detection (SCD) is essential for cultivated land protection, yet existing benchmarks and models remain insufficient for fine-gr

Latent Causal Void: Explicit Missing-Context Reconstruction for Misinformation Detection

Model ReleasesDGX agent

arXiv:2605.12156v1 Announce Type: new Abstract: Automatic misinformation detection performs well when deception is visible in what an article explicitly states. However, some misinformation articles r

LatentHDR: Decoupling Exposure from Diffusion via Conditional Latent-to-Latent Mapping for Text/Image-to-Panoramic HDR

Model ReleasesDGX agent

arXiv:2605.11115v1 Announce Type: new Abstract: High Dynamic Range (HDR) generation remains challenging for generative models, which are largely limited to low dynamic range outputs. Recent diffusionb

Learnable Multi-level Discrete Wavelet Transforms for 3D Gaussian Splatting Frequency Modulation

Model ReleasesDGX agent

arXiv:2602.14199v2 Announce Type: replace-cross Abstract: 3D Gaussian Splatting (3DGS) has emerged as a powerful approach for novel view synthesis. However, the number of Gaussian primitives often gro

Learning Compact Boolean Networks

Model ReleasesDGX agent

arXiv:2602.05830v2 Announce Type: replace-cross Abstract: Floating-point neural networks dominate modern machine learning but incur substantial inference costs, motivating emerging interest in Boolean

Learning, Fast and Slow: Towards LLMs That Adapt Continually

Model ReleasesDGX agent

arXiv:2605.12484v1 Announce Type: new Abstract: Large language models (LLMs) are trained for downstream tasks by updating their parameters (e.g., via RL). However, updating parameters forces them to a

Learning Subspace-Preserving Sparse Attention Graphs from Heterogeneous Multiview Data

Model ReleasesDGX agent

arXiv:2605.11881v1 Announce Type: new Abstract: The high-dimensional features extracted from large-scale unlabeled data via various pretrained models with diverse architectures are referred to as hete

Learning to Foresee: Unveiling the Unlocking Efficiency of On-Policy Distillation

Model ReleasesDGX agent

arXiv:2605.11739v1 Announce Type: new Abstract: On-policy distillation (OPD) has emerged as an efficient post-training paradigm for large language models. However, existing studies largely attribute t

Learning What Matters: Adaptive Information-Theoretic Objectives for Robot Exploration

Model ReleasesDGX agent

arXiv:2605.12084v1 Announce Type: cross Abstract: Designing learnable information-theoretic objectives for robot exploration remains challenging. Such objectives aim to guide exploration toward data t

Less than one hour until Falcon 9’s launch of Dragon’s 34th Commercial Resupply Services mission to the @Space_Station. Teams continue to mo…

Model ReleasesDGX agent

Less than one hour until Falcon 9’s launch of Dragon’s 34th Commercial Resupply Services mission to the @Space_Station. Teams continue to monitor weather, which is currently 10% favorable for liftoff

← Previous
1…256257258259260…377
Next →