AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

Reassessing High-Performing LLMs on Polish Medical Exams: True Competence or Bias-Driven Performance?

DGX agent

arXiv:2606.12250v1 Announce Type: new Abstract: Large language models (LLMs) in medicine are mainly evaluated using multiple-choice question answering (MCQA), which can overestimate real clinical abil

model-releasesarxiv-cs-cl
11 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

ReMoT: Reinforcement Learning with Motion Contrast Triplets

DGX agent

arXiv:2603.00461v3 Announce Type: replace Abstract: We present ReMoT, a unified training paradigm to systematically address the fundamental shortcomings of VLMs in spatio-temporal consistency -- a cri

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

Reroute, Don't Remove: Recoverable Visual Token Routing for Vision-Language Models

DGX agent

arXiv:2606.12412v1 Announce Type: cross Abstract: Vision-language models (VLMs) project images into hundreds to thousands of visual tokens, making decoder inference expensive in both attention computa

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Restless bandits with imperfect binary feedback: PCL-indexability analysis and computation

DGX agent

arXiv:2606.11192v1 Announce Type: new Abstract: We study restless bandits with binary latent states and imperfect binary feedback, motivated by opportunistic spectrum access with sensing errors. For t

model-releasesarxiv-cs-lg
11 Jun 2026
Model Releases

Robust Privacy: Inference-Stage Privacy through Certified Robustness

DGX agent

arXiv:2601.17360v2 Announce Type: replace-cross Abstract: An adversary observing a model's released prediction can infer sensitive attributes of the queried input, or even reconstruct representatives

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Robustness of Mixtures of Experts to Feature Noise

DGX agent

arXiv:2601.14792v2 Announce Type: replace Abstract: Despite their practical success, it remains unclear why Mixture of Experts (MoE) models can outperform dense networks beyond sheer parameter scaling

model-releasesarxiv-cs-lg
11 Jun 2026
Model Releases

RoVE: Rotary Value Embeddings Attention for Relative Position-dependent Value Pathways

DGX agent

arXiv:2606.11275v1 Announce Type: cross Abstract: Rotary Position Embeddings (RoPE) make attention scores position-relative but leave the value pathway position-blind: the message sent by a value toke

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

RSTR: Reducing SpatioTemporal Redundancy in Diffusion Transformers

DGX agent

arXiv:2512.14096v2 Announce Type: replace Abstract: Diffusion Transformers (DiTs) have achieved remarkable success in image generation, yet their deployment is hindered by high computational costs. We

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

Scaling Laws of Global Weather Models

DGX agent

arXiv:2602.22962v2 Announce Type: replace Abstract: Data-driven models are revolutionizing weather forecasting. To optimize training efficiency and model performance, this paper analyzes empirical sca

model-releasesarxiv-cs-lg
11 Jun 2026
Model Releases

SceneMiner: Identity-Preserving Multi-Task Fine-Tuning for Unified BEV Scene Mining

DGX agent

arXiv:2606.11507v1 Announce Type: new Abstract: Mining hard, safety-critical scenes from driving logs is bottlenecked by the absence of difficulty labels, and no single proxy, collision risk, trajecto

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

SheafStain: Sheaf-Theoretic Schrodinger Bridge for Spatially and Biologically Coherent Virtual Staining

DGX agent

arXiv:2606.11846v1 Announce Type: new Abstract: Current virtual staining approaches offer the potential for time- and cost-efficient biomarker quantification in cancer diagnostics and prognostics. How

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

Simplicity Suffices for Parameter Noise Injection in Stochastic Gradient Descent

DGX agent

arXiv:2606.12054v1 Announce Type: new Abstract: Injecting noise into the optimization process is a well-established technique for improving the training and generalization of deep neural networks. Yet

model-releasesarxiv-cs-lg
11 Jun 2026
Model Releases

SirenFNO: Efficient and Full Frequency Learning of Fourier Neural Operators

DGX agent

arXiv:2606.11518v1 Announce Type: cross Abstract: Fourier neural operators (FNOs) are effective and efficient surrogates for approximating solutions of PDEs and generalize across discretizations. Howe

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Soft-Prompt Tuning for Fair and Efficient LLM Benchmark Evaluation

DGX agent

arXiv:2606.12117v1 Announce Type: cross Abstract: Benchmark scores often misrepresent a large language model's (LLM's) knowledge, because they rely, e.g., on the model's ability to follow specific for

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

SoftMatcha 2: A Fast and Soft Pattern Matcher for Trillion-Scale Corpora

DGX agent

arXiv:2602.10908v2 Announce Type: replace Abstract: We present SoftMatcha 2, an ultra-fast and flexible search algorithm that enables search over trillion-scale natural language corpora in under 0.3 s

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

Sparse probes and murky physics: a case study of interpretability challenges in a foundation model for continuum dynamics

DGX agent

arXiv:2606.11657v1 Announce Type: cross Abstract: Generative AI emulators are increasingly used in scientific domains where we already have strong theory, benchmarks, and physical intuition. This rais

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Sparsified Kolmogorov-Arnold Networks for Interpretable Quantum State Tomography

DGX agent

arXiv:2606.11814v1 Announce Type: cross Abstract: Machine-learning approaches to quantum state tomography can achieve high reconstruction fidelity, but the physical structure used by the trained model

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Spatially Coupled Phase-to-Depth Calibration for Fringe Projection Profilometry

DGX agent

arXiv:2606.11601v1 Announce Type: new Abstract: In fringe projection profilometry (FPP), depth is commonly recovered by fitting a phase-to-depth relation independently at each camera pixel. Although s

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

SPEA2^+: Improved Density Estimation in SPEA2 with Provable Runtime Guarantees

DGX agent

arXiv:2606.12382v1 Announce Type: cross Abstract: The Strength Pareto Evolutionary Algorithm 2 (SPEA2) is a popular and prominent evolutionary algorithm for solving multi-objective optimisation proble

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

SPEAR: A System for Post-Quantization Error-Adaptive Recovery Enabling Efficient Low-Bit LLM Serving

DGX agent

arXiv:2606.11244v1 Announce Type: cross Abstract: Efficient large language model (LLM) serving is increasingly constrained by deployment cost. Quantization is a key technique for reducing serving cost

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

STEAM: Squeeze and Transform Enhanced Attention Module

DGX agent

arXiv:2412.09023v3 Announce Type: replace Abstract: Channel and spatial attention mechanisms introduced in earlier work enhance the representational capabilities of deep convolutional neural networks

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

Steering the Noise: Turning Random Perturbations into Effective Descent for Memory-Efficient LLM Fine-Tuning

DGX agent

arXiv:2601.04710v2 Announce Type: replace Abstract: Fine-tuning large language models (LLMs) achieves strong performance but is often limited by the memory overhead of backpropagation. Zeroth-order (Z

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

Substrate Asymmetry in User-Side Memory: A Diagnostic Framework

DGX agent

arXiv:2606.11712v1 Announce Type: cross Abstract: User-side memory in LLMs is typically scored as a single 'personalization' capability: given a user's history, is the output more user-aware? We show

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

SwiftCTS: Fast Cross-Design Prediction and Pareto Optimization of Clock Tree Metrics via Few-Shot Calibration

DGX agent

arXiv:2606.11348v1 Announce Type: new Abstract: Clock Tree Synthesis (CTS) is a computationally expensive stage in the physical design flow, requiring iterative EDA tool invocations to navigate a vast

model-releasesarxiv-cs-lg
11 Jun 2026
Model Releases

System Report for CCL25-Eval Task 5: New Dataset and LoRA-Fine-Tuned Qwen2.5

DGX agent

arXiv:2606.12392v1 Announce Type: cross Abstract: Recently, large language models (LLMs) have achieved promising progress in the fields of classical Chinese translation and the generation of classical

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Tac-DINO: Learning Vision-Tactile Features with Patch Alignment

DGX agent

arXiv:2606.12069v1 Announce Type: new Abstract: Touch is the primary medium through which humans interact with the environment. Currently, tactile learning mainly focuses on image-level pretraining or

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

TAHOE: Text-to-SQL with Automated Hint Optimization from Experience

DGX agent

arXiv:2606.12387v1 Announce Type: cross Abstract: Large Language Models (LLMs) have democratized database access through Text-to-SQL, but moving from prototypes to production remains difficult. Real d

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Task-Aligned Stability Analysis of Vision-Language Models for Autonomous Driving Hazard Detection

DGX agent

arXiv:2606.11889v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly used for scene understanding in autonomous driving, but robustness analysis often relies on task-agnost

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Teaching Diffusion to Speculate Left-to-Right

DGX agent

arXiv:2606.11552v1 Announce Type: new Abstract: Large language models (LLMs) achieve remarkable performance across a wide range of tasks, but their autoregressive decoding process incurs substantial i

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

The Language You Ask In: Language-Conditioned Ideological Divergence in LLM Analysis of Contested Political Documents

DGX agent

arXiv:2601.12164v3 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly deployed as analytical tools across multilingual contexts, yet their outputs may carry systemati

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

The N-Body Problem: Parallel Execution from Single-Person Egocentric Video

DGX agent

arXiv:2512.11393v2 Announce Type: replace Abstract: Humans can intuitively parallelise complex activities, but can a model predict this from observing a single person? Given one egocentric video, we i

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

The Standard Interpretable Model: A general theory of interpretable machine learning to deductively design interpretable methods using Lagrangian mechanics

DGX agent

arXiv:2606.12289v1 Announce Type: cross Abstract: As Artificial Intelligence models grow in complexity, interpretability has become an indispensable tool for understanding, debugging, and controlling

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

The Structural Attention Tax: How Retrieval Format Hijacks In-Context Learning Independent of Content

DGX agent

arXiv:2606.11198v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) systems inject external knowledge to improve LLM outputs, yet the format of injected content -- distinct from its

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Time-multiplexed layer reuse for physical neural networks

DGX agent

arXiv:2511.00044v3 Announce Type: replace Abstract: Physical neural networks (PNNs) are promising candidates for next-generation computing, but existing demonstrations remain several orders of magnitu

model-releasesarxiv-cs-lg
11 Jun 2026
Model Releases

Time-Series Foundation Model Embeddings for Remaining Useful Life Estimation

DGX agent

arXiv:2606.11990v1 Announce Type: cross Abstract: Remaining Useful Life (RUL) prediction is essential for industrial predictive maintenance, yet many learning-based approaches rely on extensive featur

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

TimeRouter: Efficient and Adaptive Routing of Time-Series Foundation Models

DGX agent

arXiv:2606.11625v1 Announce Type: new Abstract: Time-series foundation models (TSFMs) are increasingly explored as predictive experts within emerging agentic time-series systems. However, TSFMs exhibi

model-releasesarxiv-cs-lg
11 Jun 2026
Model Releases

TouchThinker: Scaling Tactile Commonsense Reasoning to the Open World with Large-scale Data and Action-aware Representation

DGX agent

arXiv:2606.11637v1 Announce Type: new Abstract: Touch is a key modality for embodied agents to understand the physical world. Although recent work has incorporated tactile signals into language system

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Toward Generalist Autonomous Research via Hypothesis-Tree Refinement

DGX agent

arXiv:2606.11926v1 Announce Type: cross Abstract: Scientific progress depends on a repeated loop of exploration, experimentation, and abstraction. Researchers test candidate directions, interpret the

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Toward Trustworthy AI: Multi-Target Adversarial Attacks and Robust Defenses for Continuous Data Summarization

DGX agent

arXiv:2606.11804v1 Announce Type: new Abstract: Trustworthy AI requires reliable data-processing pipelines, not only robust downstream predictive models. As an upstream component, data summarization d

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Towards Data-free and Training-free Compression for Speech Foundation Models Using Parameter Clustering

DGX agent

arXiv:2606.11836v1 Announce Type: cross Abstract: This paper presents a novel data-free and training-free compression approach for speech foundation models using channelwise clustering via k-means. Mo

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Towards Fully Automated Exam Grading: Fairness-Aware Recognition of Handwritten Answers with Foundation Models

DGX agent

arXiv:2606.11477v1 Announce Type: cross Abstract: Correcting handwritten exams by hand is time-consuming and error-prone, particularly for large cohorts, while fully digital exams tend to force a dida

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Using Explainability as a Training-Time Reliability Signal for Efficient ECG Classification

DGX agent

arXiv:2606.12252v1 Announce Type: cross Abstract: Training deep neural networks for clinical time-series analysis is computationally demanding, yet many healthcare settings lack the resources required

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Verifiable Environments Are LEGO Bricks: Recursive Composition for Reasoning Generalization

DGX agent

arXiv:2606.12373v1 Announce Type: new Abstract: Reinforcement Learning (RL) with verifiable environments has emerged as a powerful approach for enhancing the reasoning capabilities of Large Language M

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

VICX: Generalizable Robot Manipulation via Video Generation and In-Context Operator Network

DGX agent

arXiv:2606.12028v1 Announce Type: new Abstract: Generalizable robot manipulation requires not only task-level reasoning over unseen scenes, but also reliable grounding of visual plans into embodiment-

model-releasesarxiv-cs-ro
11 Jun 2026
Model Releases

VietMed-MCQ: A Consistency-Filtered Data Synthesis Framework for Vietnamese Traditional Medicine Evaluation

DGX agent

arXiv:2601.03792v2 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated remarkable proficiency in general medical domains. However, their performance significantly degrades

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

Visualizing LLM Latent Space Geometry Through Dimensionality Reduction

DGX agent

arXiv:2511.21594v3 Announce Type: replace Abstract: Large language models (LLMs) achieve state-of-the-art results across many natural language tasks, but their internal mechanisms remain difficult to

model-releasesarxiv-cs-lg
11 Jun 2026
Model Releases

VL-DINO: Leveraging CLIP Vision-Language Knowledge for Open-Vocabulary Object Detectio

DGX agent

arXiv:2606.11546v1 Announce Type: new Abstract: Vision-language models like CLIP can provide rich semantic priors for open-vocabulary object detection. However, jointly integrating both textual and vi

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

When Does Language Matter? Multilingual Instructions Reveal Step-wise Language Sensitivity in Vision-Language-Action Models

DGX agent

arXiv:2606.11906v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have shown strong performance in language-conditioned robotic manipulation, yet their robustness to linguistic varia

model-releasesarxiv-cs-cl
11 Jun 2026
← Previous
1…134135136137138…361
Next →