AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,131 results
Model Releases

Decoupled DiLoCo for Resilient Distributed Pre-training

DGX agent

arXiv:2604.21428v1 Announce Type: new Abstract: Modern large-scale language model pre-training relies heavily on the single program multiple data (SPMD) paradigm, which requires tight coupling across

model-releasesarxiv-cs-cl
24 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Deep FinResearch Bench: Evaluating AI's Ability to Conduct Professional Financial Investment Research

DGX agent

arXiv:2604.21006v1 Announce Type: new Abstract: We introduce Deep FinResearch Bench, a practical and comprehensive evaluation framework for deep research (DR) agents in financial investment research.

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

DenoiseRank: Learning to Rank by Diffusion Models

DGX agent

arXiv:2604.20852v1 Announce Type: cross Abstract: Learning to rank (LTR) is one of the core tasks in Machine Learning. Traditional LTR models have made great progress, but nearly all of them are imple

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Dialect vs Demographics: Quantifying LLM Bias from Implicit Linguistic Signals vs. Explicit User Profiles

DGX agent

arXiv:2604.21152v1 Announce Type: cross Abstract: As state-of-the-art Large Language Models (LLMs) have become ubiquitous, ensuring equitable performance across diverse demographics is critical. Howev

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Differentially Private Model Merging

DGX agent

arXiv:2604.20985v1 Announce Type: cross Abstract: In machine learning applications, privacy requirements during inference or deployment time could change constantly due to varying policies, regulation

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Do LLMs Overthink Basic Math Reasoning? Benchmarking the Accuracy-Efficiency Tradeoff in Language Models

DGX agent

arXiv:2507.04023v3 Announce Type: replace Abstract: Large language models (LLMs) achieve impressive performance on complex mathematical benchmarks yet sometimes fail on basic math reasoning while gene

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

Do MLLMs Understand Pointing? Benchmarking and Enhancing Referential Reasoning in Egocentric Vision

DGX agent

arXiv:2604.21461v1 Announce Type: new Abstract: Egocentric AI agents, such as smart glasses, rely on pointing gestures to resolve referential ambiguities in natural language commands. However, despite

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

Domain-Aware Hierarchical Contrastive Learning for Semi-Supervised Generalization Fault Diagnosis

DGX agent

arXiv:2604.20928v1 Announce Type: cross Abstract: Fault diagnosis under unseen operating conditions remains highly challenging when labeled data are scarce. Semi-supervised domain generalization fault

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Dr. Assistant: Enhancing Clinical Diagnostic Inquiry via Structured Diagnostic Reasoning Data and Reinforcement Learning

DGX agent

arXiv:2601.13690v2 Announce Type: replace Abstract: Clinical Decision Support Systems (CDSSs) provide reasoning and inquiry guidance for physicians, yet they face notable challenges, including high ma

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

Droplet-LNO: Physics-Informed Laplace Neural Operators for Accurate Prediction of Droplet Spreading Dynamics on Complex Surfaces

DGX agent

arXiv:2604.20993v1 Announce Type: new Abstract: Spreading of liquid droplets on solid substrates constitutes a classic multiphysics problem with widespread applications ranging from inkjet printing, s

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Drug Synergy Prediction via Residual Graph Isomorphism Networks and Attention Mechanisms

DGX agent

arXiv:2604.21473v1 Announce Type: cross Abstract: In the treatment of complex diseases, treatment regimens using a single drug often yield limited efficacy and can lead to drug resistance. In contrast

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

EARL-BO: Reinforcement Learning for Multi-Step Lookahead, High-Dimensional Bayesian Optimization

DGX agent

arXiv:2411.00171v2 Announce Type: replace Abstract: To avoid myopic behavior, multi-step lookahead Bayesian optimization (BO) algorithms consider the sequential nature of BO and have demonstrated prom

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Efficient Multi-Source Knowledge Transfer by Model Merging

DGX agent

arXiv:2508.19353v2 Announce Type: replace-cross Abstract: While transfer learning is an effective strategy, it often overlooks the opportunity to leverage knowledge from numerous available models onli

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

Empirical Comparison of Agent Communication Protocols for Task Orchestration

DGX agent

arXiv:2603.22823v3 Announce Type: replace Abstract: Context. The problem of comparative evaluation of communication protocols for task orchestration by large language model (LLM) agents is considered.

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

EngramaBench: Evaluating Long-Term Conversational Memory with Structured Graph Retrieval

DGX agent

arXiv:2604.21229v1 Announce Type: cross Abstract: Large language model assistants are increasingly expected to retain and reason over information accumulated across many sessions. We introduce Engrama

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Enhancing Science Classroom Discourse Analysis through Joint Multi-Task Learning for Reasoning-Component Classification

DGX agent

arXiv:2604.21137v1 Announce Type: cross Abstract: Analyzing the reasoning patterns of students in science classrooms is critical for understanding knowledge construction mechanism and improving instru

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Evaluating AI Meeting Summaries with a Reusable Cross-Domain Pipeline

DGX agent

arXiv:2604.21345v1 Announce Type: new Abstract: We present a reusable evaluation pipeline for generative AI applications, instantiated for AI meeting summaries and released with a public artifact pack

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

EVENT5Ws: A Large Dataset for Open-Domain Event Extraction from Documents

DGX agent

arXiv:2604.21890v1 Announce Type: new Abstract: Event extraction identifies the central aspects of events from text. It supports event understanding and analysis, which is crucial for tasks such as in

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

FairyFuse: Multiplication-Free LLM Inference on CPUs via Fused Ternary Kernels

DGX agent

arXiv:2604.20913v1 Announce Type: new Abstract: Large language models are increasingly deployed on CPU-only platforms where memory bandwidth is the primary bottleneck for autoregressive generation. We

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Fake or Real, Can Robots Tell? Evaluating VLM Robustness to Domain Shift in Single-View Robotic Scene Understanding

DGX agent

arXiv:2506.19579v3 Announce Type: replace-cross Abstract: Robotic scene understanding increasingly relies on Vision-Language Models (VLMs) to generate natural language descriptions of the environment.

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Federated Co-tuning Framework for Large and Small Language Models

DGX agent

arXiv:2411.11707v3 Announce Type: replace-cross Abstract: By adapting Large Language Models (LLMs) to domain-specific tasks or enriching them with domain-specific knowledge, we can fully harness the c

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Federated Learning for Surgical Vision in Appendicitis Classification: Results of the FedSurg EndoVis 2024 Challenge

DGX agent

arXiv:2510.04772v2 Announce Type: replace-cross Abstract: Developing generalizable surgical AI requires multi-institutional data, yet patient privacy constraints preclude direct data sharing, making F

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Fine-Tuning Regimes Define Distinct Continual Learning Problems

DGX agent

arXiv:2604.21927v1 Announce Type: new Abstract: Continual learning (CL) studies how models acquire tasks sequentially while retaining previously learned knowledge. Despite substantial progress in benc

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Flow Matching for Conditional MRI-CT and CBCT-CT Image Synthesis

DGX agent

arXiv:2510.04823v2 Announce Type: replace Abstract: Generating synthetic CT (sCT) from MRI or CBCT plays a crucial role in enabling MRI-only and CBCT-based adaptive radiotherapy, improving treatment p

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

From Codebooks to VLMs: Evaluating Automated Visual Discourse Analysis for Climate Change on Social Media

DGX agent

arXiv:2604.21786v1 Announce Type: new Abstract: Social media platforms have become primary arenas for climate communication, generating millions of images and posts that - if systematically analysed -

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

From Research Question to Scientific Workflow: Leveraging Agentic AI for Science Automation

DGX agent

arXiv:2604.21910v1 Announce Type: new Abstract: Scientific workflow systems automate execution -- scheduling, fault tolerance, resource management -- but not the semantic translation that precedes it.

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

GeCo: Evaluating Geometric Consistency for Video Generation via Motion and Structure

DGX agent

arXiv:2512.22274v2 Announce Type: replace Abstract: We introduce GeCo, a geometry-grounded metric for jointly detecting geometric deformation and occlusion-inconsistency artifacts in static scenes. By

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

Generalizing Numerical Reasoning in Table Data through Operation Sketches and Self-Supervised Learning

DGX agent

arXiv:2604.21495v1 Announce Type: cross Abstract: Numerical reasoning over expert-domain tables often exhibits high in-domain accuracy but limited robustness to domain shift. Models trained with super

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Geo-R1: Improving Few-Shot Geospatial Referring Expression Understanding with Reinforcement Fine-Tuning

DGX agent

arXiv:2509.21976v3 Announce Type: replace-cross Abstract: Referring expression understanding in remote sensing poses unique challenges, as it requires reasoning over complex object-context relationshi

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Geometric Characterisation and Structured Trajectory Surrogates for Clinical Dataset Condensation

DGX agent

arXiv:2604.21638v1 Announce Type: new Abstract: Dataset condensation constructs compact synthetic datasets that retain the training utility of large real-world datasets, enabling efficient model devel

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Geometric Monomial (GEM): a family of rational 2N-differentiable activation functions

DGX agent

arXiv:2604.21677v1 Announce Type: cross Abstract: The choice of activation function plays a crucial role in the optimization and performance of deep neural networks. While the Rectified Linear Unit (R

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

GeoMind: An Agentic Workflow for Lithology Classification with Reasoned Tool Invocation

DGX agent

arXiv:2604.21501v1 Announce Type: new Abstract: Lithology classification in well logs is a fundamental geoscience data mining task that aims to infer rock types from multi dimensional geophysical sequ

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

GeoRA: Geometry-Aware Low-Rank Adaptation for RLVR

DGX agent

arXiv:2601.09361v3 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) is a key paradigm for improving large-scale reasoning models. Unlike supervised fine-tun

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

GerAV: Towards New Heights in German Authorship Verification using Fine-Tuned LLMs on a New Benchmark

DGX agent

arXiv:2601.13711v2 Announce Type: replace Abstract: Authorship verification (AV) is the task of determining whether two texts were written by the same author and has been studied extensively, predomin

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

GiVA: Gradient-Informed Bases for Vector-Based Adaptation

DGX agent

arXiv:2604.21901v1 Announce Type: cross Abstract: As model sizes continue to grow, parameter-efficient fine-tuning has emerged as a powerful alternative to full fine-tuning. While LoRA is widely adopt

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Grounding Machine Creativity in Game Design Knowledge Representations: Empirical Probing of LLM-Based Executable Synthesis of Goal Playable Patterns under Structural Constraints

DGX agent

arXiv:2603.07101v3 Announce Type: replace Abstract: Creatively translating complex gameplay ideas into executable artifacts (e.g., games as Unity projects and code) remains a central challenge in comp

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Grounding Video Reasoning in Physical Signals

DGX agent

arXiv:2604.21873v1 Announce Type: new Abstract: Physical video understanding requires more than naming an event correctly. A model can answer a question about pouring, sliding, or collision from textu

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

HWE-Bench: Benchmarking LLM Agents on Real-World Hardware Bug Repair Tasks

DGX agent

arXiv:2604.14709v2 Announce Type: replace Abstract: Existing benchmarks for hardware design primarily evaluate Large Language Models (LLMs) on isolated, component-level tasks such as generating HDL mo

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

HyperAdapt: Simple High-Rank Adaptation

DGX agent

arXiv:2509.18629v3 Announce Type: replace-cross Abstract: Foundation models excel across diverse tasks, but adapting them to specialized applications often requires fine-tuning, an approach that is me

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

HyperFM: An Efficient Hyperspectral Foundation Model with Spectral Grouping

DGX agent

arXiv:2604.21127v1 Announce Type: new Abstract: The NASA PACE mission provides unprecedented hyperspectral observations of ocean color, aerosols, and clouds, offering new insights into how these compo

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

Hyperloop Transformers

DGX agent

arXiv:2604.21254v1 Announce Type: cross Abstract: LLM architecture research generally aims to maximize model quality subject to fixed compute/latency budgets. However, many applications of interest su

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

ICNN-enhanced 2SP: Leveraging input convex neural networks for solving two-stage stochastic programming

DGX agent

arXiv:2505.05261v3 Announce Type: replace-cross Abstract: Two-stage stochastic programming (2SP) offers a basic framework for modelling decision-making under uncertainty, yet scalability remains a cha

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Ideological Bias in LLMs' Economic Causal Reasoning

DGX agent

arXiv:2604.21334v1 Announce Type: new Abstract: Do large language models (LLMs) exhibit systematic ideological bias when reasoning about economic causal effects? As LLMs are increasingly used in polic

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

ILDR: Geometric Early Detection of Grokking

DGX agent

arXiv:2604.20923v1 Announce Type: new Abstract: Grokking describes a delayed generalization phenomenon in which a neural network achieves perfect training accuracy long before validation accuracy impr

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Intent Laundering: AI Safety Datasets Are Not What They Seem

DGX agent

arXiv:2602.16729v3 Announce Type: replace-cross Abstract: We systematically evaluate the quality of widely used adversarial safety datasets from two perspectives: in isolation and in practice. In isol

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Interpretable facial dynamics as behavioral and perceptual traces of deepfakes

DGX agent

arXiv:2604.21760v1 Announce Type: new Abstract: Deepfake detection research has largely converged on deep learning approaches that, despite strong benchmark performance, offer limited insight into wha

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

IRIS: Interpolative Renyi Iterative Self-play for Large Language Model Fine-Tuning

DGX agent

arXiv:2604.20933v1 Announce Type: cross Abstract: Self-play fine-tuning enables large language models to improve beyond supervised fine-tuning without additional human annotations by contrasting annot

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

It's High Time: A Survey of Temporal Question Answering

DGX agent

arXiv:2505.20243v4 Announce Type: replace Abstract: Time plays a critical role in how information is generated, retrieved, and interpreted. In this survey, we provide a comprehensive overview of Tempo

model-releasesarxiv-cs-cl
24 Apr 2026
← Previous
1…301302303304305…357
Next →