AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,131 results
Model Releases

MEDIC: Comprehensive Evaluation of Leading Indicators for LLM Safety and Utility in Clinical Applications

DGX agent

arXiv:2409.07314v4 Announce Type: replace Abstract: While Large Language Models (LLMs) achieve superhuman performance on standardized medical licensing exams, these static benchmarks have become satur

model-releasesarxiv-cs-cl
30 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Mergeable Model-Side Aggregation States for Long-Context Language Models

DGX agent

arXiv:2607.26448v1 Announce Type: new Abstract: A known limitation of long-context language models is their increasingly unreliable performance in non-additive, set-based aggregation as context length

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Meta-Learned Reward Shaping for Reinforcement Learning from Human Feedback

DGX agent

arXiv:2607.26094v1 Announce Type: cross Abstract: Reinforcement Learning from Human Feedback (RLHF) is the standard approach for aligning large language models with human preferences, but its quality

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Mixture-of-experts for handwriting trajectory reconstruction from IMU sensors

DGX agent

arXiv:2607.26708v1 Announce Type: new Abstract: The use of digital pens for online handwriting trajectory reconstruction is a prevalent method for human-computer interaction. In this study, we focus o

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

ML2B: Benchmarking LLMs on Cross-Lingual ML Pipeline Generation

DGX agent

arXiv:2509.22768v3 Announce Type: replace Abstract: We introduce ML2B, the first benchmark for evaluating cross-lingual task comprehension in end-to-end ML pipeline generation by large language models

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

OmegaUse-OfficeVal: Benchmarking LLM Agents on Long-Horizon Office-Suite Tasks with Economic Grounding

DGX agent

arXiv:2607.27155v1 Announce Type: cross Abstract: Large language model (LLM) agents are increasingly expected to assist users in completing tasks. However, existing benchmarks provide limited support

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

OmniAD: Detect and Understand Industrial Anomaly via Multimodal Reasoning

DGX agent

arXiv:2505.22039v2 Announce Type: replace Abstract: While anomaly detection has made significant progress, generating detailed analyses that incorporate industrial knowledge remains a challenge. To ad

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

Online Handwriting Trajectory Reconstruction from Kinematic Sensors using Temporal Convolutional Network

DGX agent

arXiv:2607.26733v1 Announce Type: new Abstract: Handwriting with digital pens is a common way to facilitate human-computer interaction through the use of Online Handwriting (OH) trajectory reconstruct

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

Parameter-Free Dynamic Regret for Online Convex Optimization under Heavy-Tailed Noise

DGX agent

arXiv:2607.27073v1 Announce Type: new Abstract: We study online convex optimization (OCO) in non-stationary environments under heavy-tailed noise, where the stochastic gradient oracle admits only a fi

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

PatchDenoiser: Parameter-efficient multi-scale patch learning and fusion denoiser for Low-dose CT imaging

DGX agent

arXiv:2602.21987v3 Announce Type: replace Abstract: Low-dose CT images are essential for reducing radiation exposure in cancer screening, pediatric imaging, and longitudinal monitoring protocols, but

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

Persistence Spheres: a Bi-continuous Linear Representation of Measures for Partial Optimal Transport

DGX agent

arXiv:2603.15384v2 Announce Type: replace-cross Abstract: We improve and extend persistence spheres, introduced in~ite{pegoraro2025persistence}. Persistence spheres map an integrable measure mu on the

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Phoneme- vs. Character-Level Targets and Selective State-Space Models for Intracortical Brain-to-Text

DGX agent

arXiv:2607.26751v1 Announce Type: new Abstract: State-of-the-art intracortical brain-to-text systems pair a neural-sequence phone decoder with an external language model. Two design axes remain undere

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Position: Evaluation Scores Are Perishable Knowledge Claims

DGX agent

arXiv:2607.26191v1 Announce Type: cross Abstract: Evaluation methodologies for language models increasingly combine multiple signals, from automated metrics and LLM-as-judge ratings to human assessmen

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Post-Training at the Edge of Detectability: A Game-Theoretic Approach to Fine-Tuning

DGX agent

arXiv:2607.26358v1 Announce Type: new Abstract: Reinforcement learning (RL) fine-tuning is widely used in language model training to improve model performance on a target task while limiting drift fro

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

PowerAtlas: Towards Electricity-Computing Co-Scheduling for Power Systems

DGX agent

arXiv:2607.26710v1 Announce Type: new Abstract: The rapid growth of AI workloads is turning data centers into large-scale, volatile, yet spatiotemporally flexible grid loads, creating an urgent need f

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Progressive Multimodal Alignment for Continual Instruction Tuning

DGX agent

arXiv:2607.26947v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) rely on a projector to align visual representations with the language embedding space, making it central to cro

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

Projective Graph Residualization: Variation-Allocation Frontiers for Control-Function IV

DGX agent

arXiv:2606.14636v2 Announce Type: replace Abstract: Control-function instrumental-variable estimators pass an estimated first-stage residual to an outcome model. The residual must retain the latent co

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Prosody-driven Jailbreaks in Audio LLMs: A Controlled Study and Mechanistic Analysis

DGX agent

arXiv:2607.26541v1 Announce Type: cross Abstract: Audio-capable foundation models enable end-to-end spoken interaction, but they also introduce safety risks beyond transcript content. It remains uncle

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Rad-JEPA 3D: Radiology Joint-Embedding Predictive Model for 3D Computed Tomography

DGX agent

arXiv:2607.26196v1 Announce Type: new Abstract: Self-supervised pretraining is central to 3D medical image analysis, where unlabeled CT volumes are abundant but expert annotations are scarce. Yet exis

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

RAGuard: A Layered Defense Framework for Retrieval-Augmented Generation Systems Against Data Poisoning

DGX agent

arXiv:2607.26339v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) systems ground large language models (LLMs) in external corpora, but this reliance exposes them to corpus poisoning

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

ReCo: Reweighting GRPO Against Distributional Concentration

DGX agent

arXiv:2607.26862v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) has become a standard reinforcement learning method for post-training language models. Recent work shows that

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Reeling It In: Flexible Needle Pick Up via Thread Manipulation for Autonomous Suturing

DGX agent

arXiv:2607.26337v1 Announce Type: new Abstract: Suture-needle pickup is necessary for autonomous suturing, as a needle can be unexpectedly dropped or strategically released to adjust the grasping conf

model-releasesarxiv-cs-ro
30 Jul 2026
Model Releases

Rethinking Self-Evolution: A Constrained Exploration-Exploitation Process for Mitigating Skill Overfitting

DGX agent

arXiv:2607.26643v1 Announce Type: cross Abstract: Enabling large language model (LLM) agents to accumulate and reuse experience from past interactions remains a central challenge in real-world applica

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Risk-Aware Motion Planning with Learned Trajectory Primitives and Probabilistic Safety Assessment

DGX agent

arXiv:2607.26802v1 Announce Type: new Abstract: This paper presents a radial basis function network (RBFN)-informed motion planning framework for safe and efficient urban autonomous driving. The propo

model-releasesarxiv-cs-ro
30 Jul 2026
Model Releases

Route by Kinematics, Act by Observation: Kinematics-Supervised Expert Routing in MoE-Augmented VLA

DGX agent

arXiv:2607.26807v1 Announce Type: new Abstract: While MoE augments VLA via expert specialization, router suffers from ineffective expert routing owing to the kinematic heterogeneity of actions across

model-releasesarxiv-cs-ro
30 Jul 2026
Model Releases

Same Evidence, Different Target: Decoding How Diagnostic Evidence Bears on Causal Questions from Language-Model States

DGX agent

arXiv:2607.26929v1 Announce Type: new Abstract: The same diagnostic result can support or challenge one causal claim yet fail to address another when the claims concern different populations, outcomes

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

SciFigQual-Bench: A Benchmark for Scientific Figure Quality Assessment with Full-Manuscript Context

DGX agent

arXiv:2607.27084v1 Announce Type: new Abstract: Scientific images are the core elements of presenting experimental conclusions, elaborating system architecture, and supporting comparative arguments in

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

SecRespond: Benchmarking AI Agents for Real-World Post-Compromise Incident Response

DGX agent

arXiv:2607.26791v1 Announce Type: cross Abstract: Large Language Model (LLM) agents are increasingly adopted in real-world security operations with access to host artifacts and command-line interfaces

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Seeing or Knowing? Visual Context Sensitivity in Multimodal Large Language Models

DGX agent

arXiv:2607.26326v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) achieve strong performance by integrating visual inputs with the rich priors of pretrained language models. How

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

Semantic-Aware Temporal Adaptation for UAV Anti-UAV Tracking

DGX agent

arXiv:2607.26511v1 Announce Type: new Abstract: UAV Anti-UAV tracking is an emerging low-altitude security task for localizing an adversarial UAV using the onboard camera of a moving observer UAV. It

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

SERPO: Self-Evolving Rubric Policy Optimization for Open-Ended Test-Time Reinforcement Learning

DGX agent

arXiv:2607.26873v1 Announce Type: new Abstract: Test-time reinforcement learning (TTRL) enables language models to self-evolve at inference time without labeled feedback. Existing methods rely on answ

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Setoka: A Benchmark for Hierarchical User Understanding in Personalized Agents over Heterogeneous Data

DGX agent

arXiv:2607.27056v1 Announce Type: cross Abstract: Personalized agents are increasingly applied to assist users across a wide range of tasks. Effective personalized assistance requires not only retriev

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Shared SFT Lessons Across Alignment, Model Organisms, and Toy Models

DGX agent

arXiv:2607.26173v1 Announce Type: new Abstract: Alignment training, model organisms, and toy models are usually treated as separate research areas. But projects in all three frequently use supervised

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Simultaneous Coverage and Efficiency Guarantee in Online Conformal Prediction

DGX agent

arXiv:2607.26577v1 Announce Type: new Abstract: Adaptive conformal inference (ACI) of Gibbs and Cand{es and its variants are the standard approach to online conformal prediction under distribution shi

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

SpatialQ: Understanding 3D Gaussian Splatting Scene Quality via Visual-based MLLM

DGX agent

arXiv:2607.26595v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) has emerged as an effective representation for novel view synthesis and 3D scene reconstruction, creating an increasing dem

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

Structurally Separated Uncertainty in Supervised Latent Variable Models

DGX agent

arXiv:2602.11219v2 Announce Type: replace Abstract: Predictive uncertainty is commonly decomposed into epistemic and aleatoric components, but standard decompositions often produce strongly correlated

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Symphony of Bias: Exploring Gender Associations with Musical Instruments in Multimodal LLMs

DGX agent

arXiv:2607.26355v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly embedded in everyday life and widely used for information seeking, raising concerns about their potential

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Temporally Centered SIGReg Improves Multi-Task LeWorldModel Learning: From Analysis to Method

DGX agent

arXiv:2607.26924v1 Announce Type: new Abstract: Recent work on LeWorldModel (LeWM) has shown that the Sketched Isotropic Gaussian Regularizer (SIGReg) enables stable end-to-end world-model learning fr

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

The Art of Not Forgetting A Local Learning Architecture for Continual Learning

DGX agent

arXiv:2607.26523v1 Announce Type: new Abstract: We introduce CMP (Cognitive Memory Primitive), a continual-learning architecture that repre?sents inputs as sparse relational codes, stores them in a tw

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

The Reliability of LLMs for Medical Diagnosis: An Examination of Consistency, Manipulation, and Contextual Awareness

DGX agent

arXiv:2503.10647v2 Announce Type: replace Abstract: This study evaluated the diagnostic reliability of two Large Language Models (LLMs), Google Gemini 2.0 Flash and OpenAI ChatGPT-4o, across three dim

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Tight Generalization Bound for AdaBoost

DGX agent

arXiv:2607.26838v1 Announce Type: new Abstract: In this paper we show that the generalization error of AdaBoost is Thetaig(frac{dln(ngamma^{2}/d)}{ngamma^2}+frac{ln(1/elta)}{n}ig), where gamma is the

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

TiPToP: A Modular Open-Vocabulary Robot Manipulation System That Plans

DGX agent

arXiv:2603.09971v2 Announce Type: replace Abstract: We present TiPToP, a modular manipulation system that integrates pretrained foundation models with a GPU-accelerated Task and Motion Planner to solv

model-releasesarxiv-cs-ro
30 Jul 2026
Model Releases

Towards Trustworthy Embodied Intelligence: A Systems Framework and Graded Trustworthiness Levels

DGX agent

arXiv:2607.26121v1 Announce Type: new Abstract: Embodied intelligence integrates learned perception and decision making with real-time computation, control, and physical interaction. Because failures

model-releasesarxiv-cs-ro
30 Jul 2026
Model Releases

ToxScreen: Detecting Whether an LLM Has Been Poisoned

DGX agent

arXiv:2607.26849v1 Announce Type: cross Abstract: As large language models (LLMs) are deployed in high-stakes domains, adversaries may poison training data to implant backdoors: hidden triggers that c

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

TPCD: Tone-Pressure Contrastive Decoding and the Label-Free Gating Bottleneck in Vision-Language Models

DGX agent

arXiv:2607.26536v1 Announce Type: new Abstract: High-pressure prompts can push vision-language models (VLMs) into unsupported commitments, such as reading illegible text, reporting indeterminate times

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation

DGX agent

arXiv:2510.00192v3 Announce Type: replace Abstract: Low-rank adaptation (LoRA) has become a widely used paradigm for parameter-efficient fine-tuning of large language models, yet its representational

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Transformers Can Learn Rules They've Never Seen: Proof of Computation Beyond Interpolation

DGX agent

arXiv:2603.17019v2 Announce Type: replace Abstract: A central question in the debate over large language models is whether transformers can learn rules they have never seen, or whether they can only i

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

TreeCCA: Canonical Correlation Analysis via Gradient-Boosted Trees

DGX agent

arXiv:2607.27027v1 Announce Type: new Abstract: Gradient-boosted trees dominate tabular machine learning, yet canonical correlation analysis has always relied on linear or neural encoders. We propose

model-releasesarxiv-cs-lg
30 Jul 2026
← Previous
1…4445464748…357
Next →