AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

Fixed-Set Robustness in Programming by Example: Example Corruption and Semantic Partition Recovery

DGX agent

arXiv:2607.01280v1 Announce Type: new Abstract: Programming-by-example systems infer programs from a small set of input-output examples. Robust PBE work usually models wrong examples as samples from a

model-releasesarxiv-cs-lg
3 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Frequency Shift Physics-Informed Extreme Learning Machine for Solving High-Frequency Partial Differential Equations

DGX agent

arXiv:2607.01694v1 Announce Type: new Abstract: Solving partial differential equations (PDEs) with high-frequency solutions remains a central challenge in physics-informed machine learning due to spec

model-releasesarxiv-cs-lg
3 Jul 2026
Model Releases

From Lab to Reality: A Practical Evaluation of Deep Learning Models and LLMs for Vulnerability Detection

DGX agent

arXiv:2512.10485v2 Announce Type: replace-cross Abstract: Vulnerability detection methods based on deep learning (DL) have shown strong performance on benchmark datasets, yet their real-world effectiv

model-releasesarxiv-cs-lg
3 Jul 2026
Model Releases

From Monolingual to Multilingual: Evaluating Mamba for ASR in South African Languages

DGX agent

arXiv:2607.01502v1 Announce Type: new Abstract: Recent advances in automatic speech recognition (ASR) have explored different sequence models, including Conformer-based models and newer state space mo

model-releasesarxiv-cs-cl
3 Jul 2026
Model Releases

Generative AI and Federated Learning for Intrusion Detection Systems: A Survey

DGX agent

arXiv:2607.01305v1 Announce Type: cross Abstract: Intrusion Detection Systems (IDSs) are essential for monitoring network traffic and identifying malicious activities in modern cyber-physical, Interne

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Generic Expert Coverage for Pruning SparseMixture-of-Experts Language Models

DGX agent

arXiv:2607.01710v1 Announce Type: new Abstract: Sparsely activated Mixture-of-Experts (MoE) language models contain substantial structured redundancy among routed experts, but pruning them without dow

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Gravity-Awareness: Deep Learning Models and LLM Simulation of Human Awareness in Altered Gravity

DGX agent

arXiv:2511.05536v2 Announce Type: replace-cross Abstract: Earth s gravity fundamentally shapes human behaviour. The brain encodes this force as an internal model of gravity, enabling the prediction an

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

GroundEval: A Deterministic Replacement for LLM-as-Judge in Stateful Agent Evaluation

DGX agent

arXiv:2606.22737v2 Announce Type: replace Abstract: Before letting an agent operate over real context, can you prove it used the right evidence? GroundEval turns that question into a deterministic tes

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

HaloGuard 1.0: An Open Weights Constitutional Classifier for Multilingual AI Safety

DGX agent

arXiv:2607.02079v1 Announce Type: new Abstract: We present HaloGuard 1.0, an open-weights implementation of the constitutional-classifier paradigm for input safety. It achieves state-of-the-art perfor

model-releasesarxiv-cs-cl
3 Jul 2026
Model Releases

HarnessX: A Composable, Adaptive, and Evolvable Agent Harness Foundry

DGX agent

arXiv:2606.14249v2 Announce Type: replace Abstract: AI agent performance depends critically on the runtime harness, comprising the prompts, tools, memory, and control flow that mediate how a model obs

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Has This Checkpoint Been Abliterated? A Two-Signal Audit and Its Failure Map

DGX agent

arXiv:2607.01854v1 Announce Type: cross Abstract: Can a platform tell, before deployment, whether an open-weight checkpoint has had its refusal mechanism stripped? Runtime guards cannot: they score ge

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

HERMES: A Multi-Granularity Labeling Substrate for Pre-training Data Mixtures

DGX agent

arXiv:2607.02266v1 Announce Type: cross Abstract: Most data-mixing methods assume the corpus has already been partitioned into groups, and the choice of those groups determines what a mixer can expres

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

HNSW with Accuracy Guarantees Using Graph Spanners -- A Technical Report

DGX agent

arXiv:2607.02338v1 Announce Type: cross Abstract: Hierarchical Navigable Small World (HNSW) graphs serve as the industry standard due to their logarithmic complexity and strong empirical performance.

model-releasesarxiv-cs-cl
3 Jul 2026
Model Releases

HULAT2 at MER-TRANS 2026: Governed Multi-Agent Simplification for Spanish Easy-to-Read Generation

DGX agent

arXiv:2607.02381v1 Announce Type: new Abstract: This paper describes the participation of HULAT2-UC3M in the Spanish track of MER-TRANS 2026, a shared task on multilingual Easy-to-Read translation. Th

model-releasesarxiv-cs-cl
3 Jul 2026
Model Releases

Human Capital, Not Model Benchmarks, Predicts Hybrid Intelligence in Forecasting

DGX agent

arXiv:2607.02467v1 Announce Type: cross Abstract: Whether pairing people with AI helps or hurts is usually reported as a single average effect. Using a real-money prediction market (Polymarket) as an

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

InduceKV: Fixed-Footprint Continual Adaptation of Multimodal LLMs via Inducing KV Memories

DGX agent

arXiv:2607.02010v1 Announce Type: new Abstract: Multimodal large language models must adapt to evolving tasks and domains, yet continual improvement under bounded deployment footprint remains difficul

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Influence of Radial Basis Activation Functions on Intelligent Controller for Robotic Manipulators

DGX agent

arXiv:2607.02167v1 Announce Type: cross Abstract: This paper presents an intelligent control framework for trajectory tracking of robotic manipulators using radial basis function (RBF) neural networks

model-releasesarxiv-cs-ro
3 Jul 2026
Model Releases

IonSense-QKG: A Quantum-Readiness Metadata Framework for Lithium-Ion Battery Dataset Discovery

DGX agent

arXiv:2607.01286v1 Announce Type: new Abstract: Public lithium-ion battery datasets are increasingly used for state-of-health estimation, remaining-useful-life prediction, anomaly detection, electroch

model-releasesarxiv-cs-lg
3 Jul 2026
Model Releases

IsoSci: A Benchmark of Isomorphic Cross-Domain Science Problems for Evaluating Reasoning versus Knowledge Retrieval in LLMs

DGX agent

arXiv:2607.01431v1 Announce Type: cross Abstract: We introduce ISOSCI, a benchmark of isomorphic cross-domain science problem pairs that separates reasoning ability from domain knowledge retrieval in

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

LACUNA: A Testbed for Evaluating Localization Precision for LLM Unlearning

DGX agent

arXiv:2607.02513v1 Announce Type: cross Abstract: LLMs memorize sensitive training data, including personally identifiable information (PII), creating a pressing need for reliable post hoc removal met

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

LearNAT: Learning NL2SQL with AST-guided Task Decomposition for Large Language Models

DGX agent

arXiv:2504.02327v2 Announce Type: replace Abstract: Natural Language to SQL (NL2SQL) aims to translate natural language queries into executable SQL statements, offering non-expert users intuitive acce

model-releasesarxiv-cs-cl
3 Jul 2026
Model Releases

Learning to Move Before Learning to Do: Task-Agnostic pretraining for VLAs

DGX agent

arXiv:2607.02466v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models are fundamentally bottlenecked by the scarcity of expert demonstrations -- triplets of observations, instructions,

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Less Data, More Security: Advancing Cybersecurity LLMs Specialization via Resource-Efficient Domain-Adaptive Continuous Pre-training with Minimal Tokens

DGX agent

arXiv:2507.02964v2 Announce Type: replace-cross Abstract: The increasing scale of AI workloads demands High-Performance Computing (HPC) infrastructure and training methodologies that are both scalable

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Liquid Latent State Dynamics for Interpretable Turbofan Degradation Modeling

DGX agent

arXiv:2607.01986v1 Announce Type: new Abstract: Multivariate time-series models for prognostics are often evaluated by point prediction accuracy, yet their internal states rarely expose a coherent deg

model-releasesarxiv-cs-lg
3 Jul 2026
Model Releases

LLMs as Teaching Assistants for Mathematics Exam Grading: Reliability, and Practical Usability

DGX agent

arXiv:2607.01247v1 Announce Type: cross Abstract: Open-ended mathematics exams are valuable because they assess reasoning, proof construction, algorithmic thinking, and communication of intermediate s

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Locality-Aware Continual Unlearning for Diffusion Models

DGX agent

arXiv:2512.02657v2 Announce Type: replace-cross Abstract: Real-world deployment of text-to-image diffusion models requires continual concept removal as new privacy, copyright, or safety obligations ar

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Mastermind: Strategy-grounded Learning for Repository-Scale Vulnerability Reproduction

DGX agent

arXiv:2607.01764v1 Announce Type: new Abstract: Repository-level vulnerability reproduction is a demanding software engineering (SE) task: an agent must inspect a codebase, infer the input grammar tha

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

MedRepBench: A Comprehensive Benchmark for Medical Report Interpretation

DGX agent

arXiv:2508.16674v2 Announce Type: replace-cross Abstract: Medical report understanding from real-world document images is essential for generating patient-facing explanations and enabling structured i

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

MedStreamBench: A Time-Aware Benchmark for Streaming and Proactive Medical Video Understanding

DGX agent

arXiv:2607.01751v1 Announce Type: cross Abstract: Existing medical video benchmarks primarily evaluate whether a model produces the correct answer, but rarely assess whether it answers at the right ti

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Meta-Benchmarks for Financial-Services LLM Evaluation

DGX agent

arXiv:2607.01740v1 Announce Type: new Abstract: Public LLM leaderboards optimise for global average performance and do not capture the specific cognitive demands of financial-services work: a model th

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

MetaTT: A Global Tensor-Train Adapter for Parameter-Efficient Fine-Tuning

DGX agent

arXiv:2506.09105v3 Announce Type: replace-cross Abstract: We present MetaTT, a Tensor Train (TT) adapter framework for fine-tuning of pre-trained transformers. MetaTT enables flexible and parameter-ef

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Mixture-of-Parallelisms: Towards Memory-Efficient Training Stack for Mixture-of-Experts Models

DGX agent

arXiv:2607.01844v1 Announce Type: cross Abstract: This paper showcases a memory-efficient training stack for Mixture-of-Experts (MoE) models. It is a training paradigm that combines and specializes va

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

MKGR: Multimodal Knowledge-Graph Representation Learning for Cold-Start Protein-Protein Interaction Prediction

DGX agent

arXiv:2607.01627v1 Announce Type: cross Abstract: Accurate protein-protein interaction (PPI) prediction is central to functional genomics, disease mechanism discovery, and drug development. A difficul

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

MMBench-Live: A Continuously Evolving Benchmark for Multimodal Models

DGX agent

arXiv:2607.01813v1 Announce Type: cross Abstract: Evaluation benchmarks are essential for assessing vision-language models (VLMs), but most multimodal benchmarks are static, making them vulnerable to

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

MMIR-TCM: Memory-Integrated Multimodal Inference and Retrieval for TCM Clinical Decision Support

DGX agent

arXiv:2607.01814v1 Announce Type: new Abstract: Traditional Chinese Medicine (TCM) diagnosis, particularly through tongue inspection, faces persistent challenges in subjectivity and reproducibility. T

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Model Merging as Probabilistic Inference in Fine-Tuning Parameter Space

DGX agent

arXiv:2607.01689v1 Announce Type: cross Abstract: Model merging aims to combine existing single-task solutions into a multi-task solution without additional data-driven fine-tuning.~Most existing appr

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

MultAttnAttrib: Training-Free Multimodal Attribution in Long Document Question Answering

DGX agent

arXiv:2607.01420v1 Announce Type: cross Abstract: As grounded QA systems are increasingly deployed in AI assistants, accurately attributing generated answers to evidence is critical for user trust and

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Multilingual Prompt Localization for Agent-as-a-Judge: Language and Backbone Sensitivity in Requirement-Level Evaluation

DGX agent

arXiv:2604.04532v2 Announce Type: replace-cross Abstract: Evaluation language is typically treated as a fixed English default in agentic code benchmarks, yet we show that changing the judge's language

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

mupscaling small models: Principled warm starts and hyperparameter transfer

DGX agent

arXiv:2602.10545v2 Announce Type: replace-cross Abstract: Modern large-scale neural networks are often trained and released in multiple sizes to accommodate diverse inference budgets. To improve effic

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

NarrativeTrack: Evaluating Entity-Centric Reasoning for Narrative Understanding

DGX agent

arXiv:2601.01095v3 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) have achieved impressive progress in vision-language reasoning, yet their ability to understand tempo

model-releasesarxiv-cs-lg
3 Jul 2026
Model Releases

Neuro-Symbolic Safety Guidance for Vision-Language-Action Models via Constrained Flow Matching

DGX agent

arXiv:2607.01378v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have demonstrated promising generalization capabilities across robotic manipulation tasks, yet their real-world depl

model-releasesarxiv-cs-ro
3 Jul 2026
Model Releases

NEUROSYMLAND: Neuro-Symbolic Landing-Site Assessment for Robust and Edge-Deployable UAV Autonomy

DGX agent

arXiv:2607.02277v1 Announce Type: new Abstract: Safe landing-site assessment in unstructured environments remains a key challenge for autonomous UAV deployment, as vision-only learning approaches ofte

model-releasesarxiv-cs-ro
3 Jul 2026
Model Releases

Office Comprehension Benchmark

DGX agent

arXiv:2607.01245v1 Announce Type: cross Abstract: We introduce Office Comprehension Bench (OCB), the first public benchmark to jointly evaluate LLM systems on Word, Excel, and PowerPoint comprehension

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

OmniGAIA: Towards Native Omni-Modal AI Agents

DGX agent

arXiv:2602.22897v3 Announce Type: replace Abstract: Human intelligence naturally intertwines omni-modal perception -- spanning vision, audio, and language -- with complex reasoning and tool usage to i

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

On the Limits of Steering Vectors for Preference-Aligned Generation

DGX agent

arXiv:2607.01802v1 Announce Type: new Abstract: Steering vectors have emerged as a promising approach to controlled text generation, offering interpretable, training-free mechanisms for shaping model

model-releasesarxiv-cs-cl
3 Jul 2026
Model Releases

On the Utility and Factual Reliability of Pruned Mixture-of-Experts Models in the Biomedical Domain

DGX agent

arXiv:2607.01444v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) models offer inference speedups via selective activation but impose substantial memory requirements because the whole network

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

One More Time: Revisiting Neural Quantum States from a Reinforcement Learning Perspective

DGX agent

arXiv:2607.02292v1 Announce Type: new Abstract: Neural quantum states (NQS) provide a flexible and scalable framework for approximating quantum many-body wavefunctions. Among NQS parameterizations, au

model-releasesarxiv-cs-lg
3 Jul 2026
Model Releases

OpenSafeIntent: Evaluating Intent-Calibrated Safe Completion Across Dual-Use Prompt Sets

DGX agent

arXiv:2607.02047v1 Announce Type: cross Abstract: Safe completion requires models to provide useful assistance without enabling harm, but this behavior is difficult to evaluate with isolated prompts.

model-releasesarxiv-cs-ai
3 Jul 2026
← Previous
1…9394959697…361
Next →