AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,131 results
Model Releases

Analysing Self-Harm Representations in Language Models: a Cross-Architecture Study

DGX agent

arXiv:2607.21988v1 Announce Type: new Abstract: Self-harm content is particularly challenging to detect using NLP techniques, and is also a high-stakes task which requires the highest accuracy to enab

model-releasesarxiv-cs-cl
27 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Be Consistent! Enhancing Robust Visual Reasoning in LVLMs with Consistency Constraints

DGX agent

arXiv:2607.21722v1 Announce Type: new Abstract: While Large Vision-Language Models (LVLMs) exhibit strong perceptual capabilities, they remain vulnerable in visual reasoning tasks. Existing benchmarks

model-releasesarxiv-cs-cv
27 Jul 2026
Model Releases

Benchmarking Fine-tuning and Retrieval Strategies for a Multimodal Language Model on the NRC Reactor Operator Licensing Examination

DGX agent

arXiv:2607.22067v1 Announce Type: new Abstract: The integration of large language models (LLMs) into the nuclear power industry requires outputs grounded in domain-specific knowledge. This study evalu

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

Bilateral Trade Under Heavy-Tailed Valuations: Minimax Regret with Infinite Variance

DGX agent

arXiv:2603.06851v3 Announce Type: replace-cross Abstract: We study contextual bilateral trade under full feedback when, conditionally on the context, trader valuations have bounded density but infinit

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

Bounding the Causal Impact of ML-assisted Decision-Making via Counterfactual Correctness

DGX agent

arXiv:2607.21806v1 Announce Type: new Abstract: Predictive machine learning (ML) models are increasingly used to aid human decision-makers across various high-risk domains such as healthcare and crimi

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

Breaking the Data Barrier in Learning Symbolic Computation: A Case Study on Variable Ordering Suggestion for Cylindrical Algebraic Decomposition

DGX agent

arXiv:2601.13731v2 Announce Type: replace-cross Abstract: Symbolic computation, powered by modern computer algebra systems, has important applications in mathematical reasoning through exact deep comp

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

CARDIAG: A Dense Segment Classification Benchmark of Deep Learning Architectures for Coronary Angiography

DGX agent

arXiv:2607.22139v1 Announce Type: new Abstract: Accurate pixel-level classification of coronary angiograms is critical for cardiovascular disease assessment, yet the field lacks standardized evaluatio

model-releasesarxiv-cs-cv
27 Jul 2026
Model Releases

CEL: Comprehensive Counterfactual Explanations Library and Benchmark

DGX agent

arXiv:2607.22045v1 Announce Type: new Abstract: Counterfactual explanations are a prominent approach in explainable artificial intelligence (xAI), providing actionable guidance on what input changes w

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

Convergence analysis of a family of Zermelo-type iterations for the Bradley--Terry model

DGX agent

arXiv:2607.22221v1 Announce Type: cross Abstract: Zermelo's algorithm is a classical method for computing the maximum likelihood estimator in the Bradley--Terry (BT) model, but its convergence can be

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

Cross-Tokenizer On-Policy Distillation via Byte-Prefix Marginalization

DGX agent

arXiv:2607.22334v1 Announce Type: cross Abstract: Open-weight language models from different families exhibit complementary capabilities, motivating their consolidation into a compact student through

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

Data Quality over Capacity: Internalizing Documents into LoRA Adapters for Closed-Book QA

DGX agent

arXiv:2607.21861v1 Announce Type: new Abstract: We study baking documents directly into the weights of a 4-bit Gemma-4-e4b model via LoRA, so a system can answer questions about a corpus closed-book:

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

DBA-Bench: A Production-Fidelity Benchmark for LLM-Based Database Operations Agents

DGX agent

arXiv:2607.22165v1 Announce Type: cross Abstract: LLM-based database agents show promise, but differing task scopes, testbeds, and metrics hinder comparison. We identify four gaps between evaluation a

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

DM3D: Dynamic Mamba via Offset-Guided Feature Resampling for Point Cloud Understanding

DGX agent

arXiv:2512.03424v4 Announce Type: replace Abstract: State Space Models (SSMs) model long token sequences of point cloud with linear complexity, but require an unordered point cloud to be serialized. E

model-releasesarxiv-cs-cv
27 Jul 2026
Model Releases

Do emulated quantum circuits change what CNNs look at? Performance and explainability comparison in medical image classification

DGX agent

arXiv:2607.21186v1 Announce Type: cross Abstract: Numerous studies have analyzed the use of hybrid quantum-classical convolutional neural networks as a promising alternative to classical deep learning

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models

DGX agent

arXiv:2607.21617v1 Announce Type: cross Abstract: Vision Language Models (VLMs) are increasingly used in place of traditional OCR pipelines for document understanding. In this paper, we show they do n

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

DWT-Fusion: A Signal-Based Framework for Training-Free LLM-Generated Text Detection

DGX agent

arXiv:2607.22026v1 Announce Type: new Abstract: Detecting LLM-generated text remains challenging under zero-shot and training-free conditions, especially when detectors must generalize across datasets

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

Dynamic Commonsense Coordination for Empathetic Response Generation

DGX agent

arXiv:2607.22136v1 Announce Type: new Abstract: Empathetic Response Generation (ERG) requires models to recognize users' emotions and generate empathetic responses. Commonsense knowledge has been show

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

Encoding Invisible Causation for Bridge Diagnostic Agents: Triple-Guided Retrieval-Augmented Fine-Tuning with QLoRA

DGX agent

arXiv:2607.21680v1 Announce Type: new Abstract: Bridge infrastructure deteriorates gradually, yet its root causes---salt intrusion, freezing, fatigue cracking, and others---remain invisible to the nak

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

Energy Manifold Natural Gradient Descent: Riemannian Optimization for Neural PDE Solvers

DGX agent

arXiv:2607.22004v1 Announce Type: new Abstract: Energy natural gradient descent (ENGD) aligns parameter updates with the curvature of an underlying function-space energy, but existing formulations ass

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

Enjoy Your Talk: A Human-Centered Benchmark for Multi-Turn Dialogue with Decoupled User Simulation, Target Modeling, and Judging

DGX agent

arXiv:2607.10428v2 Announce Type: replace Abstract: Evaluating large language models (LLMs) as multi-turn conversational partners requires probing capabilities that single-turn benchmarks miss: person

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

Enough is as good as a feast: A Comprehensive Analysis of How Reinforcement Learning Mitigates Task Conflicts in LLMs

DGX agent

arXiv:2607.22039v1 Announce Type: new Abstract: Model merging plays a crucial role in consolidating multiple specialized models into a single, unified model, especially in the era of large language mo

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

Evaluation design conditions the expert-vs-auto MeSH gap: a controlled comparison of bag-of-words and BiomedBERT on the Cohen benchmark

DGX agent

arXiv:2607.21685v1 Announce Type: new Abstract: A systematic review begins with someone reading thousands of abstracts to identify the few that are relevant, and classifiers are used to prioritise tha

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

EVL-MCoT: Enhanced Vision-Language Multi-CoT for Harmful Meme Detection

DGX agent

arXiv:2607.22016v1 Announce Type: new Abstract: MEMEs are widely used on the internet and often carry strong elements of sarcasm or irony. Understanding their hidden meanings typically requires a join

model-releasesarxiv-cs-cv
27 Jul 2026
Model Releases

Explainable quantum-compressed machine learning for complex fluid flows

DGX agent

arXiv:2607.21688v1 Announce Type: cross Abstract: Machine-learning surrogates of physical systems face a paradox: explainable models facing the challenge of expressivity to capture complex nonlinear f

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

Explicit Iteration Complexity of Exact Data-Driven Inverse Optimization for Integer Linear Programs

DGX agent

arXiv:2607.22263v1 Announce Type: cross Abstract: A data-driven inverse optimization problem (DDIOP) is the problem of estimating the objective-function parameters (weights) that explain observed opti

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

Filling Before Advancing: Capability-Gap-Driven Post-Training for Scenario-Specialized Remote Sensing MLLMs

DGX agent

arXiv:2607.22205v1 Announce Type: new Abstract: Remote sensing multimodal large language models (RS-MLLMs) have improved general aerial-image understanding. However, Earth observation applications req

model-releasesarxiv-cs-cv
27 Jul 2026
Model Releases

Flight-Ready LiDAR-Inertial Odometry for Embedded Drone Platforms

DGX agent

arXiv:2607.22145v1 Announce Type: new Abstract: Open-source LiDAR-inertial odometry (LIO) systems have achieved remarkable benchmark accuracy, yet current state-of-the-art implementations are primaril

model-releasesarxiv-cs-ro
27 Jul 2026
Model Releases

fMRI2Face: A Full-HD fMRI-Video Dataset and Geometry-Guided Neural Decoding Framework for Dynamic Human Face Reconstruction

DGX agent

arXiv:2607.22302v1 Announce Type: new Abstract: Reconstructing dynamic human faces from brain activity provides a powerful way to study how the mind perceives identity, expression, and facial motion.

model-releasesarxiv-cs-cv
27 Jul 2026
Model Releases

Forensics Adapter: Unleashing CLIP for Generalizable Face Forgery Detection

DGX agent

arXiv:2411.19715v4 Announce Type: replace Abstract: We describe Forensics Adapter, an adapter network designed to transform CLIP into an effective and generalizable face forgery detector. Although CLI

model-releasesarxiv-cs-cv
27 Jul 2026
Model Releases

GLI-AL: A Multi-Modal Glioma MRI Label Resource with Unified Anatomy-Lesion Labels

DGX agent

arXiv:2607.22135v1 Announce Type: new Abstract: Existing BraTS-GLI datasets provide a widely used benchmark for adult glioma MRI segmentation, but their task definition focuses on tumor subregions and

model-releasesarxiv-cs-cv
27 Jul 2026
Model Releases

Ground Truth First: A Longitudinal Evaluation Instrument for Agent Memory, and the Tenure Crossover in Memory-Architecture Rankings

DGX agent

arXiv:2607.21962v1 Announce Type: new Abstract: Benchmarks for LLM-agent memory typically generate conversations first and extract answer keys afterwards -- with documented label-error and contaminati

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

Hyperball May Not Be a Free Lunch

DGX agent

arXiv:2607.22444v1 Announce Type: new Abstract: For scale-invariant deep networks, Hyperball-style optimizers have shown strong performance in large-scale training by fixing the norms of matrix-valued

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

IFCLoRA: Topology-Aware Rank Allocation for Parameter-Efficient Fine-Tuning

DGX agent

arXiv:2607.22251v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) is a widely used parameter-efficient fine-tuning method for large language models, but its performance depends strongly on ho

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

Improving Large Vision-Language Models' Understanding for Flow Field Data

DGX agent

arXiv:2507.18311v3 Announce Type: replace Abstract: Large Vision-Language Models (LVLMs) have shown impressive capabilities across a range of tasks that integrate visual and textual understanding, suc

model-releasesarxiv-cs-cv
27 Jul 2026
Model Releases

Indexing: the Beginning and the End

DGX agent

arXiv:2607.22361v1 Announce Type: new Abstract: We study information bottlenecks in modern deep-learning architectures -- RNNs, softmax transformers, linear-attention transformers and state-space mode

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

InteractComp: Evaluating Search Agents With Ambiguous Queries

DGX agent

arXiv:2510.24668v2 Announce Type: replace Abstract: Language agents have demonstrated remarkable potential in web search and information retrieval. However, many search-agent benchmarks assume that us

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

Interpretable Anomaly and Drift Detection with Gaussian Mixture Models

DGX agent

arXiv:2607.16811v2 Announce Type: replace Abstract: We revisit Gaussian Mixture Models (GMMs) as a lightweight, interpretable tool for anomaly detection and, in particular, for detecting distributiona

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

Interpretable EEG biomarkers with bag-of-waves: Spatial and temporal waveform dictionaries for low-data regimes

DGX agent

arXiv:2607.22508v1 Announce Type: new Abstract: Electroencephalography (EEG) is widely used to diagnose neurological conditions, but its analysis usually relies on either predefined spectral features

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

IR275K: A Benchmark for Infrared Multi-Frame Super-Resolution Toward Efficient Remote Sensing

DGX agent

arXiv:2607.22380v1 Announce Type: new Abstract: Efficient processing is becoming increasingly important in infrared remote sensing, where satellite constellations produce large volumes of observations

model-releasesarxiv-cs-cv
27 Jul 2026
Model Releases

J-CoT: Chain-of-Thought in J-Space

DGX agent

arXiv:2607.21981v1 Announce Type: new Abstract: Chain-of-thought prompting improves language-model reasoning by carrying intermediate states across successive computation steps. However, relying on na

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

k{appa}-LoRA: Condition Numbers Reveal Which LoRA Matrices Worth Updating

DGX agent

arXiv:2607.22489v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) has become a widely adopted technique for efficient neural network fine-tuning, decomposing model updates into low-rank matri

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

Khondo: A Multimodal Benchmark for Document Packet Splitting of Bangla Forms

DGX agent

arXiv:2607.21780v1 Announce Type: new Abstract: Document packets, multiple documents concatenated into a single file, are common in government and administrative workflows, yet splitting them into the

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

Language-Aware Distillation for Multilingual Instruction-Following Speech LLMs with ASR-Only Supervision

DGX agent

arXiv:2603.07025v2 Announce Type: replace Abstract: Speech Large Language Models (LLMs) that understand and follow instructions in many languages are useful for real-world interaction, but are difficu

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

Latent PDE mapping for efficient physics-informed learning across geometries with limited data

DGX agent

arXiv:2607.22215v1 Announce Type: new Abstract: In this study, we introduce latent PDE mapping, a broadly applicable physics-informed learning technique designed to enable efficient geometric generali

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

Layer-wise LoRA fine-tuning: a similarity metric approach

DGX agent

arXiv:2602.05988v2 Announce Type: replace Abstract: Pre-training Large Language Models (LLMs) on web-scale datasets becomes fundamental for advancing general-purpose AI. In contrast, enhancing their p

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

LeAct: Learning to Reason from Expert Actions

DGX agent

arXiv:2607.21856v1 Announce Type: cross Abstract: Modern reasoning models depend on reasoning data, today sourced from human annotations or distilled from stronger LLMs. However, a rich and largely un

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

Learning What Matters: Supervising Sparse Attention Routing with Causal Evidence Sets

DGX agent

arXiv:2607.21692v1 Announce Type: cross Abstract: Sparse attention reduces the cost of long contexts by allowing each query to read only selected parts of the input. These selectors are often trained

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

LLM-Based Visual Explanation Evaluation Framework for Assessing the Explainability of Facial Skin Disease Classification Models

DGX agent

arXiv:2606.16794v2 Announce Type: replace Abstract: This study proposes a domain-specific LLM-based Visual Explanation Evaluation Framework for assessing visual attention explanations in facial skin d

model-releasesarxiv-cs-cv
27 Jul 2026
← Previous
1…5657585960…357
Next →