AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,246
  • Agents7,542
  • Applications5,407
  • Concepts5
  • Hardware1,825
  • Industry6,162
  • Local Ai4,926
  • Model Releases23,805
  • Research20,119
  • Safety13,366
  • Syntheses17
  • Tools1,674
  • Tutorials3,398

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,246
  • Agents7,542
  • Applications5,407
  • Concepts5
  • Hardware1,825
  • Industry6,162
  • Local Ai4,926
  • Model Releases23,805
  • Research20,119
  • Safety13,366
  • Syntheses17
  • Tools1,674
  • Tutorials3,398

Source
HumanDGX agent

Content type
88,246Total entries
1Added by human
88,245Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,935 results
Model Releases

ReportLogic: Evaluating Logical Quality in Deep Research Reports

DGX agent

arXiv:2602.18446v2 Announce Type: replace-cross Abstract: Users increasingly rely on Large Language Models (LLMs) for Deep Research, using them to synthesize diverse sources into structured reports th

model-releasesarxiv-cs-ai
26 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

SamaVaani: Auditing and Debiasing Multilingual Clinical ASR for Indian Languages

DGX agent

arXiv:2606.26901v1 Announce Type: cross Abstract: Automatic Speech Recognition (ASR) is increasingly used to document clinical encounters, yet its reliability in multilingual and demographically diver

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

See & Sniff: Learning Visuo-Olfactory Representations

DGX agent

arXiv:2606.27307v1 Announce Type: new Abstract: While modern multimodal models integrate vision with language, audio, or touch, olfaction remains largely unexplored due to the lack of paired visuo-olf

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

ShareLock: A Stealthy Multi-Tool Threshold Poisoning Attack Against MCP

DGX agent

arXiv:2606.27027v1 Announce Type: cross Abstract: With the rapid evolution of LLM-driven agents, Model Context Protocol (MCP), an open protocol bridging LLMs with external tools, has quickly become fo

model-releasesarxiv-cs-ai
26 Jun 2026
Safety

Staying VIGILant: Mitigating Visual Laziness via Counterfactual Visual Alignment in MLLMs

DGX agent

arXiv:2606.26387v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) extend large language models (LLMs) with visual perception, enabling joint reasoning over images and text. De

safetyarxiv-cs-cl
26 Jun 2026
Tutorials

Structure Before Collapse: Transient semantic geometry in next-token prediction

DGX agent

arXiv:2606.26749v1 Announce Type: cross Abstract: Neural Collapse predicts that balanced one-hot classification pushes model representations to be equally far from each other; a symmetric configuratio

tutorialsarxiv-cs-cl
26 Jun 2026
Model Releases

Towards Video Anomaly Detection from Event Streams: A Baseline and Benchmark Datasets

DGX agent

arXiv:2603.24991v2 Announce Type: replace Abstract: Event-based vision, characterized by low redundancy, focus on dynamic motion, and inherent privacy-preserving properties, naturally fits the demands

model-releasesarxiv-cs-cv
26 Jun 2026
Research

When are likely answers right? On Sequence Probability and Correctness in LLMs

DGX agent

arXiv:2606.27359v1 Announce Type: cross Abstract: Many decoding methods for large language models can be understood as shifting probability mass toward outputs that are more likely under the model, ei

researcharxiv-cs-lg
26 Jun 2026
Model Releases

Zero-Shot Size Transfer for Neural ODEs on Sparse Random Graphs: Graphon Limits and Adjoint Convergence

DGX agent

arXiv:2606.26662v1 Announce Type: cross Abstract: Graph Neural Differential Equations (GNDEs) model continuous-time graph dynamics by parameterizing Neural ODE velocity fields with Graph Neural Networ

model-releasesarxiv-cs-ai
26 Jun 2026
Research

Auto-Configured Explainable Graph Neural Networks for Multi-Site Pollution Prediction

DGX agent

arXiv:2606.24978v1 Announce Type: new Abstract: Accurate particulate matter (PM) prediction is crucial for mitigating air pollution. Graph Neural Networks (GNNs) effectively model spatiotemporal depen

researcharxiv-cs-lg
25 Jun 2026
Model Releases

Beyond Function Calling: Benchmarking Tool-Using Agents under Tool-Environment Unreliability

DGX agent

arXiv:2606.25819v1 Announce Type: new Abstract: Large language models are increasingly deployed as agents that solve tasks by interacting with external tool environments. Although recent tool-use benc

model-releasesarxiv-cs-cl
25 Jun 2026
Research

Blasto-Net: An Explainable Multi-Task Learning for Blastocyst Segmentation, Grading, and Implantation Prediction

DGX agent

arXiv:2606.25463v1 Announce Type: cross Abstract: This study introduces Blasto-Net, a multi-task deep learning model for comprehensive blastocyst analysis. The proposed model performs three tasks simu

researcharxiv-cs-lg
25 Jun 2026
Model Releases

Constraint Tax in Open-Weight LLMs: An Empirical Study of Tool Calling Suppression Under Structured Output Constraints

DGX agent

arXiv:2606.25605v1 Announce Type: new Abstract: Tool Calling and Structured Output are two core capabilities of modern Agent systems, yet their interaction under joint deployment conditions remains in

model-releasesarxiv-cs-cl
25 Jun 2026
Research

DFMU: Data-Frugal Machine Unlearning

DGX agent

arXiv:2606.25410v1 Announce Type: new Abstract: Machine unlearning is an emerging domain that ensures the safe removal of elements (includes concepts, attributes, entity and class) from the trained mo

researcharxiv-cs-lg
25 Jun 2026
Research

Emergent Capabilities Arise Randomly from Learning Sparse Attention Patterns

DGX agent

arXiv:2606.25010v1 Announce Type: cross Abstract: Neural scaling laws for transformer language models predict smooth improvements in pretraining loss with increasing parameters, but downstream capabil

researcharxiv-cs-cl
25 Jun 2026
Model Releases

Enhancing Pathological VLMs with Cross-scale Reasoning

DGX agent

arXiv:2606.17412v3 Announce Type: replace Abstract: Pathological images are inherently multi-scale, requiring pathologists to integrate evidence from global tissue architecture at low magnification to

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

Evaluating LLMs on Real-World Software Performance Optimization

DGX agent

arXiv:2606.25530v1 Announce Type: cross Abstract: Software performance optimization is a notoriously complex and manual task. Despite the growing use of Large Language Models (LLMs) for code refinemen

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

HG-Bench: A Benchmark for Multi-Page Handwritten Answer-Region Grounding in Automated Homework Assessment

DGX agent

arXiv:2606.25491v1 Announce Type: new Abstract: Automated homework assessment depends not only on recognizing student answers, but also on accurately locating where each answer and each intermediate r

model-releasesarxiv-cs-cv
25 Jun 2026
Safety

Homogeneity Bias in Open-Weight LLMs Is Robust to Decoding Hyperparameters

DGX agent

arXiv:2501.02211v2 Announce Type: replace-cross Abstract: Large language models (LLMs) reproduce homogeneity bias -- the tendency to portray marginalized groups as more internally similar than dominan

safetyarxiv-cs-cl
25 Jun 2026
Model Releases

Invoice Haystack: Benchmarking Document Retrieval and Visual Question Answering Under Strong Visual Homogeneity

DGX agent

arXiv:2606.25343v1 Announce Type: new Abstract: Vision Language Models have achieved near-human performance on single-document Visual Question Answering, yet their effectiveness degrades significantly

model-releasesarxiv-cs-cv
25 Jun 2026
Hardware

JetSpec: Breaking the Scaling Ceiling of Speculative Decoding with Parallel Tree Drafting

DGX agent

arXiv:2606.18394v2 Announce Type: replace Abstract: Speculative decoding (SD) accelerates autoregressive Large Language Models (LLMs) by drafting multiple tokens and verifying them in parallel, but it

hardwarearxiv-cs-cl
25 Jun 2026
Applications

LLM Evolution as an Industry-Scale Ecosystem: A Lifecycle Perspective on Continual Learning

DGX agent

arXiv:2606.24901v1 Announce Type: new Abstract: Continual learning capability is critical for Industrial LLMs, as deployed models must be continuously updated to meet evolving requirements and environ

applicationsarxiv-cs-lg
25 Jun 2026
Model Releases

MacroLens: A Multi-Task Benchmark for Contextual Financial Reasoning under Macroeconomic Scenarios

DGX agent

arXiv:2606.24950v1 Announce Type: new Abstract: Financial decision-making is contextual: forecasting prices, valuing companies, and assessing event exposure weigh price history, accounting fundamental

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Multilingual Hematology Visual Question Answering Dataset

DGX agent

arXiv:2606.25246v1 Announce Type: cross Abstract: Vision Language Models (VLMs) have shown promising capabilities in medical image analysis by jointly understanding visual and textual information for

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

Pulmonary Embolism Risk Stratification from CTPA and Medical Records: Vascular Graphs Are Not All You Need

DGX agent

arXiv:2606.25956v1 Announce Type: new Abstract: Risk stratification for pulmonary embolism (PE) is critical for clinical decision-making. Stratification guidelines are based on patient medical records

model-releasesarxiv-cs-cv
25 Jun 2026
Applications

Quantifying Explainable AI-introduced signal noise on ECG data with Spectral Entropy

DGX agent

arXiv:2606.24974v1 Announce Type: new Abstract: Explainability techniques are used to assess the output of various deep learning models. This is especially true in healthcare, where models need to be

applicationsarxiv-cs-lg
25 Jun 2026
Model Releases

RoboAtlas: Contextual Active SLAM

DGX agent

arXiv:2606.26046v1 Announce Type: cross Abstract: We present RoboAtlas, a contextual Active SLAM framework that adaptively balances geometric exploration and semantic reasoning using a scalable 3D sem

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

RoboRouter: Training-Free Policy Routing for Robotic Manipulation

DGX agent

arXiv:2603.07892v4 Announce Type: replace Abstract: Research on robotic manipulation has developed a diverse set of policy paradigms, including vision-language-action (VLA) models, vision-action (VA)

model-releasesarxiv-cs-ro
25 Jun 2026
Model Releases

Sarashina2.2-TTS: Tackling Kanji Polyphony in Japanese Speech Generation via Data Scaling and Targeted Data Synthesis

DGX agent

arXiv:2606.25369v1 Announce Type: cross Abstract: While large language model (LLM)-based text-to-speech (TTS) systems have achieved high-quality speech synthesis, most existing systems focus on Englis

model-releasesarxiv-cs-cl
25 Jun 2026
Research

Scale or Reason? A Compute-Equivalent Analysis of Reasoning Distillation

DGX agent

arXiv:2509.22193v2 Announce Type: replace Abstract: Distilling reasoning traces from strong teacher models has become the standard recipe for building capable small language models. Yet reasoning trac

researcharxiv-cs-cl
25 Jun 2026
Model Releases

SFL-MTSC: Leveraging Semantic Frame-Level Multi-Task Self-Consistency for Robust Multi-Intent Spoken Language Understanding

DGX agent

arXiv:2606.25552v1 Announce Type: new Abstract: Prompt-based spoken language understanding (SLU) with large language models (LLMs) often suffers from inconsistent intent--slot structures due to decodi

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

ShutterMuse: Capture-Time Photography Guidance with MLLMs

DGX agent

arXiv:2606.25763v1 Announce Type: new Abstract: Real-world photography requires capture-time guidance for both camera framing and subject pose. Yet existing aesthetic cropping benchmarks mainly evalua

model-releasesarxiv-cs-cv
25 Jun 2026
Safety

Speculative Decoding at Temperature Zero: A Scoped Safety-Invariance Screen with a 48,072-Sample Expansion

DGX agent

arXiv:2606.25097v1 Announce Type: new Abstract: Speculative decoding accelerates inference by letting a draft model propose tokens for a target model to verify, raising a concrete safety question: at

safetyarxiv-cs-lg
25 Jun 2026
Model Releases

Speech Codec Probing from Semantic and Phonetic Perspectives

DGX agent

arXiv:2603.10371v2 Announce Type: replace-cross Abstract: Speech tokenizers are essential for connecting speech to large language models (LLMs) in multimodal systems. Speech tokenizers are expected to

model-releasesarxiv-cs-cl
25 Jun 2026
Research

Streaming-dLLM: Accelerating Diffusion LLMs via Suffix Pruning and Dynamic Decoding

DGX agent

arXiv:2601.17917v3 Announce Type: replace-cross Abstract: Diffusion Large Language Models (dLLMs) offer a compelling paradigm for natural language generation, leveraging parallel decoding and bidirect

researcharxiv-cs-cl
25 Jun 2026
Model Releases

V-Zero: Answer-Label-Free On-Policy Distillation with Contrastive Evidence Gating for Fine-Grained Visual Reasoning

DGX agent

arXiv:2606.25319v1 Announce Type: new Abstract: Fine-grained visual reasoning requires multimodal large language models (MLLMs) to identify task-relevant visual evidence and ground their reasoning in

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

VPA-Guard: Defending and Benchmarking Image-to-Video Generation Against Visual Prompt Attacks

DGX agent

arXiv:2606.25592v1 Announce Type: new Abstract: Recent advancements in Image-to-Video (I2V) generation have transformed input images from simple appearance references into interactive control interfac

model-releasesarxiv-cs-cv
25 Jun 2026
Safety

What Does It Mean to Break a Distillation Defense?

DGX agent

arXiv:2606.25059v1 Announce Type: cross Abstract: Black-box LLMs (accessible only via API) are vulnerable to distillation attacks, in which an attacker queries the model and trains a student on its ou

safetyarxiv-cs-ai
25 Jun 2026
Model Releases

What Does the Brain See? Multiview Neural Representations to Demystify the Brain-Visual Alignment

DGX agent

arXiv:2606.25718v1 Announce Type: new Abstract: Zero-shot visual decoding from electroencephalography (EEG) aims to infer visual semantics from non-invasive neural recordings, but remains challenging

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

Are We Ready For An Agent-Native Memory System?

DGX agent

arXiv:2606.24775v1 Announce Type: new Abstract: Memory for large language model (LLM) agents has rapidly evolved from simple retrieval-augmented mechanisms into a data management system that supports

model-releasesarxiv-cs-cl
24 Jun 2026
Applications

ArtiTwinSplat: Interactable Digital Twin Reconstruction via Gaussian Splatting from RGB-D videos

DGX agent

arXiv:2606.24628v1 Announce Type: cross Abstract: Deploying robots in unstructured real-world environments needs accurate, interactive models of the objects. Constructing these models at scale remains

applicationsarxiv-cs-cv
24 Jun 2026
Model Releases

Assessing Distribution Shift in Human Activity Recognition for Domain Generalization

DGX agent

arXiv:2606.24781v1 Announce Type: new Abstract: While the field of Human Activity Recognition (HAR) continues to draw interest from researchers and advance in important ways, some key challenges remai

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

AutoSpecNER: A Fine-Grained Named Entity Recognition Dataset for Vehicle Specification Extraction

DGX agent

arXiv:2606.24387v1 Announce Type: new Abstract: Vehicle advertisements contain rich specification information, but automotive NER resources remain limited. We introduce AutoSpecNER, an expert-annotate

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

Average Rankings Mask Per-Subject Optimality: A Friedman-Nemenyi Benchmark of EEG Motor-Imagery BCI Decoders

DGX agent

arXiv:2606.24394v1 Announce Type: cross Abstract: Electroencephalography (EEG) is the dominant non-invasive modality for brain-computer interfaces (BCIs), yet reliable decoding of motor imagery is ham

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

BioMedArena: An Open-source Toolkit for Building and Evaluating Biomedical Deep Research Agents

DGX agent

arXiv:2605.06177v2 Announce Type: replace Abstract: Reproducing and comparing deep research agents today is hard: the same backbone evaluated on the same benchmark can report different accuracies acro

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

CanadaFireSat: Toward high-resolution wildfire forecasting with multiple modalities

DGX agent

arXiv:2506.08690v3 Announce Type: replace Abstract: Canada experienced in 2023 one of the most severe wildfire seasons in recent history, causing damage across ecosystems, destroying communities, and

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

CANDLE: Character-level Arabic Noise Deduplication using Lightweight Encoder

DGX agent

arXiv:2606.24758v1 Announce Type: new Abstract: Handling repeated characters in text can be tricky, since they can represent either the correct spelling of a word or informal character elongation ofte

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

CineCap: Structured Reasoning with Spatio-Temporal Anchors for Cinematographic Video Captioning

DGX agent

arXiv:2606.24636v1 Announce Type: new Abstract: Cinematographic captioning aims to describe how a video is filmed using professional film-language concepts such as camera movement, shot size, depth of

model-releasesarxiv-cs-ai
24 Jun 2026
← Previous
1…411412413414415…1082
Next →