AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,338
  • Agents7,718
  • Applications5,508
  • Concepts5
  • Hardware1,906
  • Industry6,195
  • Local Ai5,052
  • Model Releases24,549
  • Research20,616
  • Safety13,637
  • Syntheses17
  • Tools1,678
  • Tutorials3,457

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,338
  • Agents7,718
  • Applications5,508
  • Concepts5
  • Hardware1,906
  • Industry6,195
  • Local Ai5,052
  • Model Releases24,549
  • Research20,616
  • Safety13,637
  • Syntheses17
  • Tools1,678
  • Tutorials3,457

Source
HumanDGX agent

Content type
90,338Total entries
1Added by human
90,337Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
53,226 results
Model Releases

AeTHERON: Autoregressive Topology-aware Heterogeneous Graph Operator Network for Fluid-Structure Interaction

DGX agent

arXiv:2604.13369v1 Announce Type: cross Abstract: Surrogate modeling of body-driven fluid flows where immersed moving boundaries couple structural dynamics to chaotic, unsteady fluid phenomena remains

model-releasesarxiv-cs-lg
16 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

AI Powered Image Analysis for Phishing Detection

DGX agent

arXiv:2604.13555v1 Announce Type: new Abstract: Phishing websites now rely heavily on visual imitation-copied logos, similar layouts, and matching colours-to avoid detection by text- and URL-based sys

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Analog Optical Inference on Million-Record Mortgage Data

DGX agent

arXiv:2604.13251v1 Announce Type: new Abstract: Analog optical computers promise large efficiency gains for machine learning inference, yet no demonstration has moved beyond small-scale image benchmar

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

AudioX: A Unified Framework for Anything-to-Audio Generation

DGX agent

arXiv:2503.10522v4 Announce Type: replace-cross Abstract: Audio and music generation based on flexible multimodal control signals is a widely applicable topic, with the following key challenges: 1) a

model-releasesarxiv-cs-cv
16 Apr 2026
Safety

CLIP Architecture for Abdominal CT Image-Text Alignment and Zero-Shot Learning: Investigating Batch Composition and Data Scaling

DGX agent

arXiv:2604.13561v1 Announce Type: new Abstract: Vision-language models trained with contrastive learning on paired medical images and reports show strong zero-shot diagnostic capabilities, yet the eff

safetyarxiv-cs-cv
16 Apr 2026
Model Releases

Coherence in the brain unfolds across separable temporal regimes

DGX agent

arXiv:2512.20481v4 Announce Type: replace-cross Abstract: To maintain coherence in language, the brain must satisfy key competing temporal demands: the gradual accumulation of meaning across extended

model-releasesarxiv-cs-cl
16 Apr 2026
Research

Concrete Jungle: Towards Concreteness Paved Contrastive Negative Mining for Compositional Understanding

DGX agent

arXiv:2604.13313v1 Announce Type: new Abstract: Vision-Language Models demonstrate remarkable capabilities but often struggle with compositional reasoning, exhibiting vulnerabilities regarding word or

researcharxiv-cs-lg
16 Apr 2026
Safety

Context Sensitivity Improves Human-Machine Visual Alignment

DGX agent

arXiv:2604.13883v1 Announce Type: new Abstract: Modern machine learning models typically represent inputs as fixed points in a high-dimensional embedding space. While this approach has been proven pow

safetyarxiv-cs-cv
16 Apr 2026
Model Releases

Correct Chains, Wrong Answers: Dissociating Reasoning from Output in LLM Logic

DGX agent

arXiv:2604.13065v1 Announce Type: new Abstract: LLMs can execute every step of chain-of-thought reasoning correctly and still produce wrong final answers. We introduce the Novel Operator Test, a bench

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

DeEscalWild: A Real-World Benchmark for Automated De-Escalation Training with SLMs

DGX agent

arXiv:2604.13075v1 Announce Type: new Abstract: Effective de-escalation is critical for law enforcement safety and community trust, yet traditional training methods lack scalability and realism. While

model-releasesarxiv-cs-cl
16 Apr 2026
Research

DiffMagicFace: Identity Consistent Facial Editing of Real Videos

DGX agent

arXiv:2604.13841v1 Announce Type: new Abstract: Text-conditioned image editing has greatly benefitted from the advancements in Image Diffusion Models. However, extending these techniques to facial vid

researcharxiv-cs-cv
16 Apr 2026
Research

DiT as Real-Time Rerenderer: Streaming Video Stylization with Autoregressive Diffusion Transformer

DGX agent

arXiv:2604.13509v1 Announce Type: new Abstract: Recent advances in video generation models has significantly accelerated video generation and related downstream tasks. Among these, video stylization h

researcharxiv-cs-cv
16 Apr 2026
Research

Enhancing Mixture-of-Experts Specialization via Cluster-Aware Upcycling

DGX agent

arXiv:2604.13508v1 Announce Type: new Abstract: Sparse Upcycling provides an efficient way to initialize a Mixture-of-Experts (MoE) model from pretrained dense weights instead of training from scratch

researcharxiv-cs-cv
16 Apr 2026
Model Releases

Evaluating LLM-Based Translation of a Low-Resource Technical Language: The Medical and Philosophical Greek of Galen

DGX agent

arXiv:2602.24119v2 Announce Type: replace Abstract: Purpose: This study evaluates the quality of commercial large language model (LLM) machine translation (MT) for Ancient Greek technical prose and be

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

EVE: A Domain-Specific LLM Framework for Earth Intelligence

DGX agent

arXiv:2604.13071v1 Announce Type: new Abstract: We introduce Earth Virtual Expert (EVE), the first open-source, end-to-end initiative for developing and deploying domain-specialized LLMs for Earth Int

model-releasesarxiv-cs-cl
16 Apr 2026
Applications

HiVLA: A Visual-Grounded-Centric Hierarchical Embodied Manipulation System

DGX agent

arXiv:2604.14125v1 Announce Type: new Abstract: While end-to-end Vision-Language-Action (VLA) models offer a promising paradigm for robotic manipulation, fine-tuning them on narrow control data often

applicationsarxiv-cs-cv
16 Apr 2026
Model Releases

MAny: Merge Anything for Multimodal Continual Instruction Tuning

DGX agent

arXiv:2604.14016v1 Announce Type: new Abstract: Multimodal Continual Instruction Tuning (MCIT) is essential for sequential task adaptation of Multimodal Large Language Models (MLLMs) but is severely r

model-releasesarxiv-cs-lg
16 Apr 2026
Safety

Med-CAM: Minimal Evidence for Explaining Medical Decision Making

DGX agent

arXiv:2604.13695v1 Announce Type: new Abstract: Reliable and interpretable decision-making is essential in medical imaging, where diagnostic outcomes directly influence patient care. Despite advances

safetyarxiv-cs-cv
16 Apr 2026
Model Releases

MERRIN: A Benchmark for Multimodal Evidence Retrieval and Reasoning in Noisy Web Environments

DGX agent

arXiv:2604.13418v1 Announce Type: new Abstract: Motivated by the underspecified, multi-hop nature of search queries and the multimodal, heterogeneous, and often conflicting nature of real-world web re

model-releasesarxiv-cs-cl
16 Apr 2026
Local Ai

One Token per Highly Selective Frame: Towards Extreme Compression for Long Video Understanding

DGX agent

arXiv:2604.14149v1 Announce Type: new Abstract: Long video understanding is inherently challenging for vision-language models (VLMs) because of the extensive number of frames. With each video frame ty

local-aiarxiv-cs-cv
16 Apr 2026
Research

OneHOI: Unifying Human-Object Interaction Generation and Editing

DGX agent

arXiv:2604.14062v1 Announce Type: new Abstract: Human-Object Interaction (HOI) modelling captures how humans act upon and relate to objects, typically expressed as triplets. Existing approaches split

researcharxiv-cs-cv
16 Apr 2026
Model Releases

OPTED: Open Preprocessed Trachoma Eye Dataset Using Zero-Shot SAM 3 Segmentation

DGX agent

arXiv:2603.06885v2 Announce Type: replace Abstract: Trachoma remains the leading infectious cause of blindness worldwide, with Sub-Saharan Africa bearing over 85% of the global burden and Ethiopia alo

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Out of Context: Reliability in Multimodal Anomaly Detection Requires Contextual Inference

DGX agent

arXiv:2604.13252v1 Announce Type: new Abstract: Anomaly detection aims to identify observations that deviate from expected behavior. Because anomalous events are inherently sparse, most frameworks are

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Red Skills or Blue Skills? A Dive Into Skills Published on ClawHub

DGX agent

arXiv:2604.13064v1 Announce Type: new Abstract: Skill ecosystems have emerged as an increasingly important layer in Large Language Model (LLM) agent systems, enabling reusable task packaging, public d

model-releasesarxiv-cs-cl
16 Apr 2026
Local Ai

Rhetorical Questions in LLM Representations: A Linear Probing Study

DGX agent

arXiv:2604.14128v1 Announce Type: new Abstract: Rhetorical questions are asked not to seek information but to persuade or signal stance. How large language models internally represent them remains unc

local-aiarxiv-cs-cl
16 Apr 2026
Research

Robust Ultra Low-Bit Post-Training Quantization via Stable Diagonal Curvature Estimate

DGX agent

arXiv:2604.13806v1 Announce Type: new Abstract: Large Language Models (LLMs) are widely used across many domains, but their scale makes deployment challenging. Post-Training Quantization (PTQ) reduces

researcharxiv-cs-lg
16 Apr 2026
Model Releases

RPS: Information Elicitation with Reinforcement Prompt Selection

DGX agent

arXiv:2604.13817v1 Announce Type: new Abstract: Large language models (LLMs) have shown remarkable capabilities in dialogue generation and reasoning, yet their effectiveness in eliciting user-known bu

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

SparseBalance: Load-Balanced Long Context Training with Dynamic Sparse Attention

DGX agent

arXiv:2604.13847v1 Announce Type: new Abstract: While sparse attention mitigates the computational bottleneck of long-context LLM training, its distributed training process exhibits extreme heterogene

model-releasesarxiv-cs-lg
16 Apr 2026
Research

Stability Principle Underlying Passive Dynamic Walking of Rimless Wheel

DGX agent

arXiv:2604.13530v1 Announce Type: new Abstract: Rimless wheels are known as the simplest model for passive dynamic walking. It is known that the passive gait generated only by gravity effect always be

researcharxiv-cs-ro
16 Apr 2026
Model Releases

Syn-TurnTurk: A Synthetic Dataset for Turn-Taking Prediction in Turkish Dialogues

DGX agent

arXiv:2604.13620v1 Announce Type: new Abstract: Managing natural dialogue timing is a significant challenge for voice-based chatbots. Most current systems usually rely on simple silence detection, whi

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Towards Generalizable Robotic Manipulation in Dynamic Environments

DGX agent

arXiv:2603.15620v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models excel in static manipulation but struggle in dynamic environments with moving targets. This performance gap prim

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

VGGT-Segmentor: Geometry-Enhanced Cross-View Segmentation

DGX agent

arXiv:2604.13596v1 Announce Type: new Abstract: Instance-level object segmentation across disparate egocentric and exocentric views is a fundamental challenge in visual understanding, critical for app

model-releasesarxiv-cs-cv
16 Apr 2026
Local Ai

A Hybrid Architecture for Benign-Malignant Classification of Mammography ROIs

DGX agent

arXiv:2604.12437v1 Announce Type: new Abstract: Accurate characterization of suspicious breast lesions in mammography is important for early diagnosis and treatment planning. While Convolutional Neura

local-aiarxiv-cs-cv
15 Apr 2026
Model Releases

Beyond Relevance: On the Relationship Between Retrieval and RAG Information Coverage

DGX agent

arXiv:2603.08819v3 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) systems combine document retrieval with a generative model to address complex information seeking tasks l

model-releasesarxiv-cs-ai
15 Apr 2026
Research

Calibrated Confidence Estimation for Tabular Question Answering

DGX agent

arXiv:2604.12491v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed for tabular question answering, yet calibration on structured data is largely unstudied. This pap

researcharxiv-cs-cl
15 Apr 2026
Research

Decidable By Construction: Design-Time Verification for Trustworthy AI

DGX agent

arXiv:2603.25414v2 Announce Type: replace-cross Abstract: A prevailing assumption in machine learning is that model correctness must be enforced after the fact. We observe that the properties determin

researcharxiv-cs-ai
15 Apr 2026
Research

DeepTest Tool Competition 2026: Benchmarking an LLM-Based Automotive Assistant

DGX agent

arXiv:2604.12615v1 Announce Type: new Abstract: This report summarizes the results of the first edition of the Large Language Model (LLM) Testing competition, held as part of the DeepTest workshop at

researcharxiv-cs-ai
15 Apr 2026
Research

Distorted or Fabricated? A Survey on Hallucination in Video LLMs

DGX agent

arXiv:2604.12944v1 Announce Type: cross Abstract: Despite significant progress in video-language modeling, hallucinations remain a persistent challenge in Video Large Language Models (Vid-LLMs), refer

researcharxiv-cs-ai
15 Apr 2026
Model Releases

Do VLMs Truly 'Read' Candlesticks? A Multi-Scale Benchmark for Visual Stock Price Forecasting

DGX agent

arXiv:2604.12659v1 Announce Type: cross Abstract: Vision-language models(VLMs) are increasingly applied to visual stock price forecasting, yet existing benchmarks inadequately evaluate their understan

model-releasesarxiv-cs-cl
15 Apr 2026
Model Releases

DPC-VQA: Decoupling Quality Perception and Residual Calibration for Video Quality Assessment

DGX agent

arXiv:2604.12813v1 Announce Type: new Abstract: Recent multimodal large language models (MLLMs) have shown promising performance on video quality assessment (VQA) tasks. However, adapting them to new

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

From Plan to Action: How Well Do Agents Follow the Plan?

DGX agent

arXiv:2604.12147v1 Announce Type: cross Abstract: Agents aspire to eliminate the need for task-specific prompt crafting through autonomous reason-act-observe loops. Still, they are commonly instructed

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Generative Anonymization in Event Streams

DGX agent

arXiv:2604.12803v1 Announce Type: new Abstract: Neuromorphic vision sensors offer low latency and high dynamic range, but their deployment in public spaces raises severe data protection concerns. Rece

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

GF-Score: Certified Class-Conditional Robustness Evaluation with Fairness Guarantees

DGX agent

arXiv:2604.12757v1 Announce Type: cross Abstract: Adversarial robustness is essential for deploying neural networks in safety-critical applications, yet standard evaluation methods either require expe

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Identity as Attractor: Geometric Evidence for Persistent Agent Architecture in LLM Activation Space

DGX agent

arXiv:2604.12016v1 Announce Type: new Abstract: Large language models map semantically related prompts to similar internal representations -- a phenomenon interpretable as attractor-like dynamics. We

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

INDOTABVQA: A Benchmark for Cross-Lingual Table Understanding in Bahasa Indonesia Documents

DGX agent

arXiv:2604.11970v1 Announce Type: cross Abstract: We introduce INDOTABVQA, a benchmark for evaluating cross-lingual Table Visual Question Answering (VQA) on real-world document images in Bahasa Indone

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

INFORM-CT: INtegrating LLMs and VLMs FOR Incidental Findings Management in Abdominal CT

DGX agent

arXiv:2512.14732v2 Announce Type: replace-cross Abstract: Incidental findings in CT scans, though often benign, can have significant clinical implications and should be reported following established

model-releasesarxiv-cs-ai
15 Apr 2026
Safety

Information-Geometric Decomposition of Generalization Error in Unsupervised Learning

DGX agent

arXiv:2604.12340v1 Announce Type: cross Abstract: We decompose the Kullback--Leibler generalization error (GE) -- the expected KL divergence from the data distribution to the trained model -- of unsup

safetyarxiv-cs-lg
15 Apr 2026
Model Releases

KCLarity at SemEval-2026 Task 6: Encoder and Zero-Shot Approaches to Political Evasion Detection

DGX agent

arXiv:2603.06552v2 Announce Type: replace Abstract: This paper describes the KCLarity team's participation in CLARITY, a shared task at SemEval 2026 on classifying ambiguity and evasion techniques in

model-releasesarxiv-cs-cl
15 Apr 2026
← Previous
1…459460461462463…1109
Next →