AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,620 results
Model Releases

Phi-Actor-Critic: Steering General-Sum Games to Pareto-Efficient Correlated Equilibria

DGX agent

arXiv:2606.11284v1 Announce Type: cross Abstract: Real-world multi-agent systems, from traffic coordination to resource allocation, are often modeled as general-sum games where individual incentives c

model-releasesarxiv-cs-lg
11 Jun 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Physically Constrained Ensemble Gaussian Process Modelling for Expensive Quantum Systems with Heteroskedastic Noise

DGX agent

arXiv:2606.11240v1 Announce Type: cross Abstract: Accurate modeling of quantum many-body systems often requires computationally expensive simulations such as Density Matrix Renormalization Group (DMRG

model-releasesarxiv-cs-lg
11 Jun 2026
Model Releases

PLUME: Probabilistic Latent Unified World Modeling and Parameter Estimation for Multi-Finger Manipulation

DGX agent

arXiv:2606.11396v1 Announce Type: new Abstract: Dexterous manipulation with multi-finger hands can be sensitive to physical parameters such as object shape, pose, and friction coefficients. While simu

model-releasesarxiv-cs-ro
11 Jun 2026
Model Releases

Pop quiz, which of these is no longer true (or at least directionally true), two years later?

DGX agent

Pop quiz, which of these is no longer true (or at least directionally true), two years later? 9 reasons that OpenAI could someday be seen as the WeWork of AI: 👉 Lots of competitors are catching up. 👉

model-releasesgary-marcus--x
11 Jun 2026
Model Releases

Precision-Aware Illumination-Disentangled Vision Transformer for Spacecraft 6D Pose Estimation

DGX agent

arXiv:2606.11619v1 Announce Type: new Abstract: Vision sensors provide a lightweight solution for spacecraft proximity operations, but monocular spacecraft 6D pose estimation remains difficult under i

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

ProGRank: Probe-Gradient Reranking to Defend Dense-Retriever RAG from Corpus Poisoning

DGX agent

arXiv:2603.22934v3 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) improves large language model applications by grounding generation in retrieved evidence, but also introduces c

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Q-Fold: Query-Aware Focus-Context Spatio-Temporal Folding for Long Video Understanding

DGX agent

arXiv:2606.12125v1 Announce Type: new Abstract: Long-video understanding remains challenging for multimodal large language models, because temporally extended videos often contain thousands of frames

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

Quantifying Subliminal Behavioral Transfer Ratios in Language Model Distillation

DGX agent

arXiv:2606.11270v1 Announce Type: cross Abstract: Distillation of a language model intended to transfer benign behavior to a student model may also transfer undesirable characteristics, if they are pr

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

RAIL: Rethinking Auditory Intelligence in Large Audio-Language Models with a CHC-Grounded Benchmark

DGX agent

arXiv:2606.11260v1 Announce Type: cross Abstract: Humans process rich auditory environments through tightly integrated cognitive capabilities such as audio perception, audio reasoning, and memory. Des

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Range-Aware Bayesian Optimization for Discovering Diverse Designs within Target Property Windows

DGX agent

arXiv:2606.11574v1 Announce Type: new Abstract: In many materials and product design problems, desirable candidates exhibit properties that fall within an acceptable range rather than achieve a single

model-releasesarxiv-cs-lg
11 Jun 2026
Model Releases

RankVR: Low-Rank Structure Perception and Value Recalibration for Robust Composed Image Retrieval

DGX agent

arXiv:2606.11689v1 Announce Type: new Abstract: Composed Image Retrieval (CIR) constitutes a pivotal paradigm requiring models to perform joint reasoning on reference images and modification texts. Ho

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

Reassessing High-Performing LLMs on Polish Medical Exams: True Competence or Bias-Driven Performance?

DGX agent

arXiv:2606.12250v1 Announce Type: new Abstract: Large language models (LLMs) in medicine are mainly evaluated using multiple-choice question answering (MCQA), which can overestimate real clinical abil

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

ReMoT: Reinforcement Learning with Motion Contrast Triplets

DGX agent

arXiv:2603.00461v3 Announce Type: replace Abstract: We present ReMoT, a unified training paradigm to systematically address the fundamental shortcomings of VLMs in spatio-temporal consistency -- a cri

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

Reroute, Don't Remove: Recoverable Visual Token Routing for Vision-Language Models

DGX agent

arXiv:2606.12412v1 Announce Type: cross Abstract: Vision-language models (VLMs) project images into hundreds to thousands of visual tokens, making decoder inference expensive in both attention computa

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Restless bandits with imperfect binary feedback: PCL-indexability analysis and computation

DGX agent

arXiv:2606.11192v1 Announce Type: new Abstract: We study restless bandits with binary latent states and imperfect binary feedback, motivated by opportunistic spectrum access with sensing errors. For t

model-releasesarxiv-cs-lg
11 Jun 2026
Model Releases

Robust Privacy: Inference-Stage Privacy through Certified Robustness

DGX agent

arXiv:2601.17360v2 Announce Type: replace-cross Abstract: An adversary observing a model's released prediction can infer sensitive attributes of the queried input, or even reconstruct representatives

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Robustness of Mixtures of Experts to Feature Noise

DGX agent

arXiv:2601.14792v2 Announce Type: replace Abstract: Despite their practical success, it remains unclear why Mixture of Experts (MoE) models can outperform dense networks beyond sheer parameter scaling

model-releasesarxiv-cs-lg
11 Jun 2026
Model Releases

RoVE: Rotary Value Embeddings Attention for Relative Position-dependent Value Pathways

DGX agent

arXiv:2606.11275v1 Announce Type: cross Abstract: Rotary Position Embeddings (RoPE) make attention scores position-relative but leave the value pathway position-blind: the message sent by a value toke

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

RSTR: Reducing SpatioTemporal Redundancy in Diffusion Transformers

DGX agent

arXiv:2512.14096v2 Announce Type: replace Abstract: Diffusion Transformers (DiTs) have achieved remarkable success in image generation, yet their deployment is hindered by high computational costs. We

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

Running Gemma 4 QAT 12B on an 8GB GPU at 16k context — measured the KV-cache tradeoffs

DGX agent

This post discusses running Google's Gemma 4 QAT (Quantized Aware Training) 12B model on a GPU with 8GB of memory while maintaining a 16k token context window. The author likely shares performance ben

model-releasesr-ollama
11 Jun 2026
Model Releases

Scaling Laws of Global Weather Models

DGX agent

arXiv:2602.22962v2 Announce Type: replace Abstract: Data-driven models are revolutionizing weather forecasting. To optimize training efficiency and model performance, this paper analyzes empirical sca

model-releasesarxiv-cs-lg
11 Jun 2026
Model Releases

SceneMiner: Identity-Preserving Multi-Task Fine-Tuning for Unified BEV Scene Mining

DGX agent

arXiv:2606.11507v1 Announce Type: new Abstract: Mining hard, safety-critical scenes from driving logs is bottlenecked by the absence of difficulty labels, and no single proxy, collision risk, trajecto

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

SheafStain: Sheaf-Theoretic Schrodinger Bridge for Spatially and Biologically Coherent Virtual Staining

DGX agent

arXiv:2606.11846v1 Announce Type: new Abstract: Current virtual staining approaches offer the potential for time- and cost-efficient biomarker quantification in cancer diagnostics and prognostics. How

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

Simplicity Suffices for Parameter Noise Injection in Stochastic Gradient Descent

DGX agent

arXiv:2606.12054v1 Announce Type: new Abstract: Injecting noise into the optimization process is a well-established technique for improving the training and generalization of deep neural networks. Yet

model-releasesarxiv-cs-lg
11 Jun 2026
Model Releases

SirenFNO: Efficient and Full Frequency Learning of Fourier Neural Operators

DGX agent

arXiv:2606.11518v1 Announce Type: cross Abstract: Fourier neural operators (FNOs) are effective and efficient surrogates for approximating solutions of PDEs and generalize across discretizations. Howe

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Soft-Prompt Tuning for Fair and Efficient LLM Benchmark Evaluation

DGX agent

arXiv:2606.12117v1 Announce Type: cross Abstract: Benchmark scores often misrepresent a large language model's (LLM's) knowledge, because they rely, e.g., on the model's ability to follow specific for

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

SoftMatcha 2: A Fast and Soft Pattern Matcher for Trillion-Scale Corpora

DGX agent

arXiv:2602.10908v2 Announce Type: replace Abstract: We present SoftMatcha 2, an ultra-fast and flexible search algorithm that enables search over trillion-scale natural language corpora in under 0.3 s

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

Sparse probes and murky physics: a case study of interpretability challenges in a foundation model for continuum dynamics

DGX agent

arXiv:2606.11657v1 Announce Type: cross Abstract: Generative AI emulators are increasingly used in scientific domains where we already have strong theory, benchmarks, and physical intuition. This rais

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Sparsified Kolmogorov-Arnold Networks for Interpretable Quantum State Tomography

DGX agent

arXiv:2606.11814v1 Announce Type: cross Abstract: Machine-learning approaches to quantum state tomography can achieve high reconstruction fidelity, but the physical structure used by the trained model

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Spatially Coupled Phase-to-Depth Calibration for Fringe Projection Profilometry

DGX agent

arXiv:2606.11601v1 Announce Type: new Abstract: In fringe projection profilometry (FPP), depth is commonly recovered by fitting a phase-to-depth relation independently at each camera pixel. Although s

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

SPEA2^+: Improved Density Estimation in SPEA2 with Provable Runtime Guarantees

DGX agent

arXiv:2606.12382v1 Announce Type: cross Abstract: The Strength Pareto Evolutionary Algorithm 2 (SPEA2) is a popular and prominent evolutionary algorithm for solving multi-objective optimisation proble

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

SPEAR: A System for Post-Quantization Error-Adaptive Recovery Enabling Efficient Low-Bit LLM Serving

DGX agent

arXiv:2606.11244v1 Announce Type: cross Abstract: Efficient large language model (LLM) serving is increasingly constrained by deployment cost. Quantization is a key technique for reducing serving cost

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

STEAM: Squeeze and Transform Enhanced Attention Module

DGX agent

arXiv:2412.09023v3 Announce Type: replace Abstract: Channel and spatial attention mechanisms introduced in earlier work enhance the representational capabilities of deep convolutional neural networks

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

Steering the Noise: Turning Random Perturbations into Effective Descent for Memory-Efficient LLM Fine-Tuning

DGX agent

arXiv:2601.04710v2 Announce Type: replace Abstract: Fine-tuning large language models (LLMs) achieves strong performance but is often limited by the memory overhead of backpropagation. Zeroth-order (Z

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

Substrate Asymmetry in User-Side Memory: A Diagnostic Framework

DGX agent

arXiv:2606.11712v1 Announce Type: cross Abstract: User-side memory in LLMs is typically scored as a single 'personalization' capability: given a user's history, is the output more user-aware? We show

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

SwiftCTS: Fast Cross-Design Prediction and Pareto Optimization of Clock Tree Metrics via Few-Shot Calibration

DGX agent

arXiv:2606.11348v1 Announce Type: new Abstract: Clock Tree Synthesis (CTS) is a computationally expensive stage in the physical design flow, requiring iterative EDA tool invocations to navigate a vast

model-releasesarxiv-cs-lg
11 Jun 2026
Model Releases

System Report for CCL25-Eval Task 5: New Dataset and LoRA-Fine-Tuned Qwen2.5

DGX agent

arXiv:2606.12392v1 Announce Type: cross Abstract: Recently, large language models (LLMs) have achieved promising progress in the fields of classical Chinese translation and the generation of classical

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Tac-DINO: Learning Vision-Tactile Features with Patch Alignment

DGX agent

arXiv:2606.12069v1 Announce Type: new Abstract: Touch is the primary medium through which humans interact with the environment. Currently, tactile learning mainly focuses on image-level pretraining or

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

TAHOE: Text-to-SQL with Automated Hint Optimization from Experience

DGX agent

arXiv:2606.12387v1 Announce Type: cross Abstract: Large Language Models (LLMs) have democratized database access through Text-to-SQL, but moving from prototypes to production remains difficult. Real d

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Task-Aligned Stability Analysis of Vision-Language Models for Autonomous Driving Hazard Detection

DGX agent

arXiv:2606.11889v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly used for scene understanding in autonomous driving, but robustness analysis often relies on task-agnost

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Teaching Diffusion to Speculate Left-to-Right

DGX agent

arXiv:2606.11552v1 Announce Type: new Abstract: Large language models (LLMs) achieve remarkable performance across a wide range of tasks, but their autoregressive decoding process incurs substantial i

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

The Language You Ask In: Language-Conditioned Ideological Divergence in LLM Analysis of Contested Political Documents

DGX agent

arXiv:2601.12164v3 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly deployed as analytical tools across multilingual contexts, yet their outputs may carry systemati

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

The model subsidies will eventually end and this workflow of “creating loops that will prompt your agents” will result in massive amounts of…

DGX agent

The model subsidies will eventually end and this workflow of “creating loops that will prompt your agents” will result in massive amounts of code that’s not well understood that you will have to pay l

model-releasesjerry-liu--x
11 Jun 2026
Model Releases

The N-Body Problem: Parallel Execution from Single-Person Egocentric Video

DGX agent

arXiv:2512.11393v2 Announce Type: replace Abstract: Humans can intuitively parallelise complex activities, but can a model predict this from observing a single person? Given one egocentric video, we i

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

The Standard Interpretable Model: A general theory of interpretable machine learning to deductively design interpretable methods using Lagrangian mechanics

DGX agent

arXiv:2606.12289v1 Announce Type: cross Abstract: As Artificial Intelligence models grow in complexity, interpretability has become an indispensable tool for understanding, debugging, and controlling

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

The Structural Attention Tax: How Retrieval Format Hijacks In-Context Learning Independent of Content

DGX agent

arXiv:2606.11198v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) systems inject external knowledge to improve LLM outputs, yet the format of injected content -- distinct from its

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Time-multiplexed layer reuse for physical neural networks

DGX agent

arXiv:2511.00044v3 Announce Type: replace Abstract: Physical neural networks (PNNs) are promising candidates for next-generation computing, but existing demonstrations remain several orders of magnitu

model-releasesarxiv-cs-lg
11 Jun 2026
Model Releases

Time-Series Foundation Model Embeddings for Remaining Useful Life Estimation

DGX agent

arXiv:2606.11990v1 Announce Type: cross Abstract: Remaining Useful Life (RUL) prediction is essential for industrial predictive maintenance, yet many learning-based approaches rely on extensive featur

model-releasesarxiv-cs-ai
11 Jun 2026
← Previous
1…186187188189190…472
Next →