AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,595 results
11 Jun 2026

MSUE: Multi-Modal Soccer Understanding Expert

Model ReleasesDGX agent

arXiv:2606.12106v1 Announce Type: cross Abstract: This paper presents our solution to the 2026 SoccerNet VQA Challenge. We first develop a cost-effective data synthesis pipeline driven by a Vision-Lan

Multi-Agent Reasoning with Adaptive Worker Allocation for Stance Detection

Model ReleasesDGX agent

arXiv:2606.11609v1 Announce Type: new Abstract: Stance detection requires identifying an author's position toward a target, often from short-form texts where stance is implicit, indirect, or rhetorica

Multi-View In-Cabin Monitoring System for Public Transport Vehicles

Model ReleasesDGX agent

arXiv:2606.11739v1 Announce Type: cross Abstract: We introduce a multi-view in-cabin monitoring dataset for public transportation with synchronized RGB and depth images from four inward-facing cameras


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Natural-Language Temporal Grounding in Hour-Long Videos is a Search Problem: A Benchmark and Empirical Decomposition

Model ReleasesDGX agent

arXiv:2606.12300v1 Announce Type: cross Abstract: Temporal grounding--returning the interval [t_s, t_e] for a natural-language query over a video--is the language interface to long-form video, yet has

NetBurst: Event-Centric Forecasting of Bursty, Intermittent Time Series

Model ReleasesDGX agent

arXiv:2510.22397v2 Announce Type: replace-cross Abstract: Network operators monitor their infrastructure by collecting telemetry data such as packet counts, byte rates, or flow volumes, yet answering

Neural ensemble Kalman filter: Data assimilation for compressible flows with shocks

Model ReleasesDGX agent

arXiv:2602.23461v2 Announce Type: replace-cross Abstract: Data assimilation (DA) for compressible flows with shocks is challenging because many classical DA methods generate spurious oscillations and

Neural-Parameterized Cellular Automata for Wildfire Spread

Model ReleasesDGX agent

arXiv:2606.11676v1 Announce Type: cross Abstract: Traditional wildfire models rely on rigid, low-dimensional parameters and static fuel maps, frequently underpredicting fire spread. To address this we

NightFeats @ MMU-RAGent NeurIPS 2025: A Context-Optimized Multi-Agent RAG System for the Text-to-Text Track

Model ReleasesDGX agent

arXiv:2606.11199v1 Announce Type: cross Abstract: We present NightFeats, a structured multi-agent retrieval-augmented generation (RAG) system submitted to the MMU-RAGent competition at NeurIPS 2025, w

OCSVM-Guided Representation Learning for Unsupervised Anomaly Detection

Model ReleasesDGX agent

arXiv:2507.21164v2 Announce Type: replace-cross Abstract: Unsupervised anomaly detection (UAD) aims to detect anomalies without labeled data, a necessity in many machine learning applications where an

OmniLoc: A Geometry-Aware Foundation Model for Anchor-Free UE Localization Across Diverse Indoor Environments

Model ReleasesDGX agent

arXiv:2606.11490v1 Announce Type: new Abstract: Indoor localization from wireless measurements remains challenging in large-scale deployments due to substantial variation in building geometry, the set

On Aligning Hierarchical Standardized Embedding for Audio-visual Generalized Zero-shot Learning

Model ReleasesDGX agent

arXiv:2606.11602v1 Announce Type: new Abstract: Audio-visual Generalized Zero-shot Learning (AV-GZSL) is a challenging task that aims to classify both seen and unseen objects or scenes by integrating

On the Limits of LLM-as-Judge for Scientific Novelty Assessment

Model ReleasesDGX agent

arXiv:2606.12071v1 Announce Type: cross Abstract: LLMs are increasingly used to generate and judge scientific ideas. This makes novelty evaluation a central problem. Full idea evaluation is difficult

Online Shift Detection and Conformal Adaptation for Deployed Safety Classifiers

Model ReleasesDGX agent

arXiv:2606.11949v1 Announce Type: new Abstract: We present an online monitoring system for distributional shift in deployed safety classifiers, using calibrated sequential statistics to detect when a

OpenMedReason: Scientific Reasoning Supervision for Medical Vision-Language Models

Model ReleasesDGX agent

arXiv:2606.12169v1 Announce Type: cross Abstract: High-stakes clinical use of large vision-language models (LVLMs) requires reasoning that is grounded in visual evidence and clinical knowledge, not ju

OSCS-SupCon: Orthogonal Sigmoid-based Common and Style Supervised Contrastive Learning for Robust Feature Disentanglement

Model ReleasesDGX agent

arXiv:2606.11233v1 Announce Type: new Abstract: Supervised Contrastive Learning (SupCon) has achieved strong performance by explicitly modeling pairwise relationships among samples. However, existing

Overcoming State Inertia in Full-Duplex Spoken Language Models via Activation Steering

Model ReleasesDGX agent

arXiv:2606.11386v1 Announce Type: cross Abstract: Full-duplex spoken language models (FD-SLMs) enable seamless speech interaction by allowing models to listen and speak simultaneously, yet the interna

Parameter-Efficient Adapter Tuning for Tabular-Image Multimodal Learning

Model ReleasesDGX agent

arXiv:2606.11682v1 Announce Type: new Abstract: Tabular-image multimodal learning aims to improve predictive modeling by jointly using structured tabular attributes and visual data. Although pretraine

ParseFixer: An Agentic Framework for Document Parsing via Selective Multimodal Correction

Model ReleasesDGX agent

arXiv:2606.11977v1 Announce Type: new Abstract: In this report, we present our third-place solution for the DataMFM Challenge Track 1: Document Parsing. This track requires models to recover structure

Pass@K Policy Optimization: Solving Harder Reinforcement Learning Problems

Model ReleasesDGX agent

arXiv:2505.15201v5 Announce Type: replace-cross Abstract: Reinforcement Learning (RL) algorithms sample multiple n>1 solution attempts for each problem and reward them independently. This optimizes fo

PCS-UQ: Uncertainty Quantification via the Predictability-Computability-Stability Framework

Model ReleasesDGX agent

arXiv:2505.08784v2 Announce Type: replace-cross Abstract: As machine learning (ML) enters high-stakes domains, trustworthy uncertainty quantification (UQ) is essential for safety. In this paper we int

Periodic-MAE: Periodic Video Masked Autoencoder for rPPG Estimation

Model ReleasesDGX agent

arXiv:2506.21855v2 Announce Type: replace Abstract: In this paper, we propose Periodic-MAE, a self-supervised framework for learning generalizable spatio-temporal representations of periodic physiolog

PermDoRA -- Understanding Adapter Interference in Language Models: Limits of Parameter-Space Geometry

Model ReleasesDGX agent

arXiv:2606.11262v1 Announce Type: cross Abstract: Access control in large language models (LLMs) requires modular mechanisms to enable domain-specific behavior without retraining or cross-domain inter

Phase Transitions in Attention: A Bayesian Theory of Copy Head Emergence

Model ReleasesDGX agent

arXiv:2606.12058v1 Announce Type: cross Abstract: Attention is the key mechanism underlying in-context learning in transformers, and attention patterns have been observed empirically to emerge abruptl

Phi-Actor-Critic: Steering General-Sum Games to Pareto-Efficient Correlated Equilibria

Model ReleasesDGX agent

arXiv:2606.11284v1 Announce Type: cross Abstract: Real-world multi-agent systems, from traffic coordination to resource allocation, are often modeled as general-sum games where individual incentives c

Physically Constrained Ensemble Gaussian Process Modelling for Expensive Quantum Systems with Heteroskedastic Noise

Model ReleasesDGX agent

arXiv:2606.11240v1 Announce Type: cross Abstract: Accurate modeling of quantum many-body systems often requires computationally expensive simulations such as Density Matrix Renormalization Group (DMRG

PLUME: Probabilistic Latent Unified World Modeling and Parameter Estimation for Multi-Finger Manipulation

Model ReleasesDGX agent

arXiv:2606.11396v1 Announce Type: new Abstract: Dexterous manipulation with multi-finger hands can be sensitive to physical parameters such as object shape, pose, and friction coefficients. While simu

Pop quiz, which of these is no longer true (or at least directionally true), two years later?

Model ReleasesDGX agent

Pop quiz, which of these is no longer true (or at least directionally true), two years later? 9 reasons that OpenAI could someday be seen as the WeWork of AI: 👉 Lots of competitors are catching up. 👉

Precision-Aware Illumination-Disentangled Vision Transformer for Spacecraft 6D Pose Estimation

Model ReleasesDGX agent

arXiv:2606.11619v1 Announce Type: new Abstract: Vision sensors provide a lightweight solution for spacecraft proximity operations, but monocular spacecraft 6D pose estimation remains difficult under i

ProGRank: Probe-Gradient Reranking to Defend Dense-Retriever RAG from Corpus Poisoning

Model ReleasesDGX agent

arXiv:2603.22934v3 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) improves large language model applications by grounding generation in retrieved evidence, but also introduces c

Q-Fold: Query-Aware Focus-Context Spatio-Temporal Folding for Long Video Understanding

Model ReleasesDGX agent

arXiv:2606.12125v1 Announce Type: new Abstract: Long-video understanding remains challenging for multimodal large language models, because temporally extended videos often contain thousands of frames

Quantifying Subliminal Behavioral Transfer Ratios in Language Model Distillation

Model ReleasesDGX agent

arXiv:2606.11270v1 Announce Type: cross Abstract: Distillation of a language model intended to transfer benign behavior to a student model may also transfer undesirable characteristics, if they are pr

RAIL: Rethinking Auditory Intelligence in Large Audio-Language Models with a CHC-Grounded Benchmark

Model ReleasesDGX agent

arXiv:2606.11260v1 Announce Type: cross Abstract: Humans process rich auditory environments through tightly integrated cognitive capabilities such as audio perception, audio reasoning, and memory. Des

Range-Aware Bayesian Optimization for Discovering Diverse Designs within Target Property Windows

Model ReleasesDGX agent

arXiv:2606.11574v1 Announce Type: new Abstract: In many materials and product design problems, desirable candidates exhibit properties that fall within an acceptable range rather than achieve a single

RankVR: Low-Rank Structure Perception and Value Recalibration for Robust Composed Image Retrieval

Model ReleasesDGX agent

arXiv:2606.11689v1 Announce Type: new Abstract: Composed Image Retrieval (CIR) constitutes a pivotal paradigm requiring models to perform joint reasoning on reference images and modification texts. Ho

Reassessing High-Performing LLMs on Polish Medical Exams: True Competence or Bias-Driven Performance?

Model ReleasesDGX agent

arXiv:2606.12250v1 Announce Type: new Abstract: Large language models (LLMs) in medicine are mainly evaluated using multiple-choice question answering (MCQA), which can overestimate real clinical abil

ReMoT: Reinforcement Learning with Motion Contrast Triplets

Model ReleasesDGX agent

arXiv:2603.00461v3 Announce Type: replace Abstract: We present ReMoT, a unified training paradigm to systematically address the fundamental shortcomings of VLMs in spatio-temporal consistency -- a cri

Reroute, Don't Remove: Recoverable Visual Token Routing for Vision-Language Models

Model ReleasesDGX agent

arXiv:2606.12412v1 Announce Type: cross Abstract: Vision-language models (VLMs) project images into hundreds to thousands of visual tokens, making decoder inference expensive in both attention computa

Restless bandits with imperfect binary feedback: PCL-indexability analysis and computation

Model ReleasesDGX agent

arXiv:2606.11192v1 Announce Type: new Abstract: We study restless bandits with binary latent states and imperfect binary feedback, motivated by opportunistic spectrum access with sensing errors. For t

Robust Privacy: Inference-Stage Privacy through Certified Robustness

Model ReleasesDGX agent

arXiv:2601.17360v2 Announce Type: replace-cross Abstract: An adversary observing a model's released prediction can infer sensitive attributes of the queried input, or even reconstruct representatives

Robustness of Mixtures of Experts to Feature Noise

Model ReleasesDGX agent

arXiv:2601.14792v2 Announce Type: replace Abstract: Despite their practical success, it remains unclear why Mixture of Experts (MoE) models can outperform dense networks beyond sheer parameter scaling

RoVE: Rotary Value Embeddings Attention for Relative Position-dependent Value Pathways

Model ReleasesDGX agent

arXiv:2606.11275v1 Announce Type: cross Abstract: Rotary Position Embeddings (RoPE) make attention scores position-relative but leave the value pathway position-blind: the message sent by a value toke

RSTR: Reducing SpatioTemporal Redundancy in Diffusion Transformers

Model ReleasesDGX agent

arXiv:2512.14096v2 Announce Type: replace Abstract: Diffusion Transformers (DiTs) have achieved remarkable success in image generation, yet their deployment is hindered by high computational costs. We

Running Gemma 4 QAT 12B on an 8GB GPU at 16k context — measured the KV-cache tradeoffs

Model ReleasesDGX agent

This post discusses running Google's Gemma 4 QAT (Quantized Aware Training) 12B model on a GPU with 8GB of memory while maintaining a 16k token context window. The author likely shares performance ben

Scaling Laws of Global Weather Models

Model ReleasesDGX agent

arXiv:2602.22962v2 Announce Type: replace Abstract: Data-driven models are revolutionizing weather forecasting. To optimize training efficiency and model performance, this paper analyzes empirical sca

SceneMiner: Identity-Preserving Multi-Task Fine-Tuning for Unified BEV Scene Mining

Model ReleasesDGX agent

arXiv:2606.11507v1 Announce Type: new Abstract: Mining hard, safety-critical scenes from driving logs is bottlenecked by the absence of difficulty labels, and no single proxy, collision risk, trajecto

SheafStain: Sheaf-Theoretic Schrodinger Bridge for Spatially and Biologically Coherent Virtual Staining

Model ReleasesDGX agent

arXiv:2606.11846v1 Announce Type: new Abstract: Current virtual staining approaches offer the potential for time- and cost-efficient biomarker quantification in cancer diagnostics and prognostics. How

Simplicity Suffices for Parameter Noise Injection in Stochastic Gradient Descent

Model ReleasesDGX agent

arXiv:2606.12054v1 Announce Type: new Abstract: Injecting noise into the optimization process is a well-established technique for improving the training and generalization of deep neural networks. Yet

SirenFNO: Efficient and Full Frequency Learning of Fourier Neural Operators

Model ReleasesDGX agent

arXiv:2606.11518v1 Announce Type: cross Abstract: Fourier neural operators (FNOs) are effective and efficient surrogates for approximating solutions of PDEs and generalize across discretizations. Howe

Soft-Prompt Tuning for Fair and Efficient LLM Benchmark Evaluation

Model ReleasesDGX agent

arXiv:2606.12117v1 Announce Type: cross Abstract: Benchmark scores often misrepresent a large language model's (LLM's) knowledge, because they rely, e.g., on the model's ability to follow specific for

SoftMatcha 2: A Fast and Soft Pattern Matcher for Trillion-Scale Corpora

Model ReleasesDGX agent

arXiv:2602.10908v2 Announce Type: replace Abstract: We present SoftMatcha 2, an ultra-fast and flexible search algorithm that enables search over trillion-scale natural language corpora in under 0.3 s

Sparse probes and murky physics: a case study of interpretability challenges in a foundation model for continuum dynamics

Model ReleasesDGX agent

arXiv:2606.11657v1 Announce Type: cross Abstract: Generative AI emulators are increasingly used in scientific domains where we already have strong theory, benchmarks, and physical intuition. This rais

Sparsified Kolmogorov-Arnold Networks for Interpretable Quantum State Tomography

Model ReleasesDGX agent

arXiv:2606.11814v1 Announce Type: cross Abstract: Machine-learning approaches to quantum state tomography can achieve high reconstruction fidelity, but the physical structure used by the trained model

Spatially Coupled Phase-to-Depth Calibration for Fringe Projection Profilometry

Model ReleasesDGX agent

arXiv:2606.11601v1 Announce Type: new Abstract: In fringe projection profilometry (FPP), depth is commonly recovered by fitting a phase-to-depth relation independently at each camera pixel. Although s

SPEA2^+: Improved Density Estimation in SPEA2 with Provable Runtime Guarantees

Model ReleasesDGX agent

arXiv:2606.12382v1 Announce Type: cross Abstract: The Strength Pareto Evolutionary Algorithm 2 (SPEA2) is a popular and prominent evolutionary algorithm for solving multi-objective optimisation proble

SPEAR: A System for Post-Quantization Error-Adaptive Recovery Enabling Efficient Low-Bit LLM Serving

Model ReleasesDGX agent

arXiv:2606.11244v1 Announce Type: cross Abstract: Efficient large language model (LLM) serving is increasingly constrained by deployment cost. Quantization is a key technique for reducing serving cost

STEAM: Squeeze and Transform Enhanced Attention Module

Model ReleasesDGX agent

arXiv:2412.09023v3 Announce Type: replace Abstract: Channel and spatial attention mechanisms introduced in earlier work enhance the representational capabilities of deep convolutional neural networks

Steering the Noise: Turning Random Perturbations into Effective Descent for Memory-Efficient LLM Fine-Tuning

Model ReleasesDGX agent

arXiv:2601.04710v2 Announce Type: replace Abstract: Fine-tuning large language models (LLMs) achieves strong performance but is often limited by the memory overhead of backpropagation. Zeroth-order (Z

Substrate Asymmetry in User-Side Memory: A Diagnostic Framework

Model ReleasesDGX agent

arXiv:2606.11712v1 Announce Type: cross Abstract: User-side memory in LLMs is typically scored as a single 'personalization' capability: given a user's history, is the output more user-aware? We show

SwiftCTS: Fast Cross-Design Prediction and Pareto Optimization of Clock Tree Metrics via Few-Shot Calibration

Model ReleasesDGX agent

arXiv:2606.11348v1 Announce Type: new Abstract: Clock Tree Synthesis (CTS) is a computationally expensive stage in the physical design flow, requiring iterative EDA tool invocations to navigate a vast

System Report for CCL25-Eval Task 5: New Dataset and LoRA-Fine-Tuned Qwen2.5

Model ReleasesDGX agent

arXiv:2606.12392v1 Announce Type: cross Abstract: Recently, large language models (LLMs) have achieved promising progress in the fields of classical Chinese translation and the generation of classical

← Previous
1…147148149150151…377
Next →