AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,315 results
21 Apr 2026

AWPD: Frequency Shield Network for Agnostic Watermark Presence Detection

Model ReleasesDGX agent

arXiv:2603.06723v3 Announce Type: replace Abstract: Invisible watermarks, as an essential technology for image copyright protection, have been widely deployed with the rapid development of social medi

Back to Repair: A Minimal Denoising Network for Time Series Anomaly Detection

Model ReleasesDGX agent

arXiv:2604.17388v1 Announce Type: new Abstract: We introduce JuRe (Just Repair), a minimal denoising network for time series anomaly detection that exposes a central finding: architectural complexity

Balanced Co-Clustering of Users and Items for Embedding Table Compression in Recommender Systems

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.18351v1 Announce Type: cross Abstract: Recommender systems have advanced markedly over the past decade by transforming each user/item into a dense embedding vector with deep learning models

BasketHAR: A Multimodal Dataset for Human Activity Recognition and Sport Analysis in Basketball Training Scenarios

Model ReleasesDGX agent

arXiv:2604.17065v1 Announce Type: new Abstract: Human Activity Recognition (HAR) involves the automatic identification of user activities and has gained significant research interest due to its broad

> be Yann LeCun > spend years building JEPA at Meta > company focuses on LLaMA instead > his idea stays complicated and unused > robotics pl…

Model ReleasesDGX agent

> be Yann LeCun > spend years building JEPA at Meta > company focuses on LLaMA instead > his idea stays complicated and unused > robotics plans get dropped > decides to leave and start AMI Labs > buil

BEFT: Bias-Efficient Fine-Tuning of Language Models in Low-Data Regimes

Model ReleasesDGX agent

arXiv:2509.15974v2 Announce Type: replace Abstract: Fine-tuning the bias terms of large language models (LLMs) has the potential to achieve unprecedented parameter efficiency while maintaining competi

BenchMarker: An Education-Inspired Toolkit for Highlighting Flaws in Multiple-Choice Benchmarks

Model ReleasesDGX agent

arXiv:2602.06221v2 Announce Type: replace Abstract: Multiple-choice question answering (MCQA) is standard in NLP, but benchmarks lack rigorous quality control. We present BenchMarker, an education-ins

Benchmarking Real-Time Question Answering via Executable Code Workflows

Model ReleasesDGX agent

arXiv:2604.16349v1 Announce Type: cross Abstract: Retrieving real-time information is a fundamental capability for search-integrated agents in real-world applications. However, existing benchmarks are

Benchmarking System Dynamics AI Assistants: Cloud Versus Local LLMs on CLD Extraction and Discussion

Model ReleasesDGX agent

arXiv:2604.18566v1 Announce Type: cross Abstract: We present a systematic evaluation of large language model families -- spanning both proprietary cloud APIs and locally-hosted open-source models -- o

BengaliMoralBench: A Benchmark for Auditing Moral Reasoning in Large Language Models within Bengali Language and Culture

Model ReleasesDGX agent

arXiv:2511.03180v2 Announce Type: replace Abstract: As multilingual Large Language Models (LLMs) gain traction across South Asia, their alignment with local ethical norms, particularly for Bengali, sp

Beyond Feature Fusion: Contextual Bayesian PEFT for Multimodal Uncertainty Estimation

Model ReleasesDGX agent

arXiv:2604.16615v1 Announce Type: new Abstract: We introduce CoCo-LoRA, a multimodal, uncertainty-aware parameter-efficient fine-tuning method for text prediction tasks accompanied by audio context. E

Beyond 'I Don't Know': Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty

Model ReleasesDGX agent

arXiv:2604.17293v1 Announce Type: new Abstract: Reliable Large Language Models (LLMs) should abstain when confidence is insufficient. However, prior studies often treat refusal as a generic 'I don't k

Beyond Pattern Matching: Seven Cross-Domain Techniques for Prompt Injection Detection

Model ReleasesDGX agent

arXiv:2604.18248v1 Announce Type: cross Abstract: Current open-source prompt-injection detectors converge on two architectural choices: regular-expression pattern matching and fine-tuned transformer c

Beyond Reproduction: A Paired-Task Framework for Assessing LLM Comprehension and Creativity in Literary Translation

Model ReleasesDGX agent

arXiv:2604.18169v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for creative tasks such as literary translation. Yet translational creativity remains underexplored a

Beyond Word Boundaries: A Hebrew Coreference Benchmark and an Evaluation Protocol for Morphologically Complex Text

Model ReleasesDGX agent

arXiv:2604.17108v1 Announce Type: new Abstract: Coreference Resolution (CR) is a fundamental NLP task critical for long-form tasks as information extraction, summarization, and many business applicati

Bi-LoRA: Efficient Sharpness-Aware Minimization for Fine-Tuning Large-Scale Models

Model ReleasesDGX agent

arXiv:2508.19564v2 Announce Type: replace Abstract: Fine-tuning large-scale pre-trained models with limited data presents significant challenges for generalization. While Sharpness-Aware Minimization

Bielik Guard: Efficient Polish Language Safety Classifiers for LLM Content Moderation

Model ReleasesDGX agent

arXiv:2602.07954v4 Announce Type: replace Abstract: As Large Language Models (LLMs) become increasingly deployed in Polish language applications, the need for efficient and accurate content safety cla

BOOKAGENT: Orchestrating Safety-Aware Visual Narratives via Multi-Agent Cognitive Calibration

Model ReleasesDGX agent

arXiv:2604.16541v1 Announce Type: new Abstract: Recent advancements in Large Generative Models (LGMs) have revolutionized multi-modal generation. However, generating illustrated storybooks remains an

BOP-ASK: Object-Interaction Reasoning for Vision-Language Models

Model ReleasesDGX agent

arXiv:2511.16857v3 Announce Type: replace Abstract: Vision Language Models (VLMs) have achieved impressive performance on spatial reasoning benchmarks, yet these evaluations mask critical weaknesses i

Brain-Inspired Capture: Evidence-Driven Neuromimetic Perceptual Simulation for Visual Decoding

Model ReleasesDGX agent

arXiv:2604.17927v1 Announce Type: new Abstract: Visual decoding of neurophysiological signals is a critical challenge for brain-computer interfaces (BCIs) and computational neuroscience. However, curr

BridgeEQA: Virtual Embodied Agents for Real Bridge Inspections

Model ReleasesDGX agent

arXiv:2511.12676v2 Announce Type: replace Abstract: Deploying embodied agents that can answer questions about their surroundings in realistic real-world settings remains difficult, partly due to the s

Bridging the Reasoning Gap in Vietnamese with Small Language Models via Test-Time Scaling

Model ReleasesDGX agent

arXiv:2604.17794v1 Announce Type: new Abstract: The democratization of ubiquitous AI hinges on deploying sophisticated reasoning capabilities on resource-constrained devices. However, Small Language M

Budget-Aware Anytime Reasoning with LLM-Synthesized Preference Data

Model ReleasesDGX agent

arXiv:2601.11038v2 Announce Type: replace Abstract: We study the reasoning behavior of large language models (LLMs) under limited computation budgets. In such settings, producing useful partial soluti

CAM3DNet: Comprehensively mining the multi-scale features for 3D Object Detection with Multi-View Cameras

Model ReleasesDGX agent

arXiv:2604.17024v1 Announce Type: new Abstract: Query-based 3D object detection methods using multi-view images often struggle to efficiently leverage dynamic multi-scale information, e.g., the relati

Camo-M3FD: A New Benchmark Dataset for Cross-Spectral Camouflaged Pedestrian Detection

Model ReleasesDGX agent

arXiv:2604.16582v1 Announce Type: new Abstract: Pedestrian detection is fundamental to autonomous driving, robotics, and surveillance. Despite progress in deep learning, reliable identification remain

Can Large Language Models Understand Context?

Model ReleasesDGX agent

Understanding context is key to understanding human language, an ability which Large Language Models (LLMs) have been increasingly seen to demonstrate to an impressive extent. However, though the eval

Can LLM-Generated Text Empower Surgical Vision-Language Pre-training?

Model ReleasesDGX agent

arXiv:2604.18134v1 Announce Type: new Abstract: Recent advancements in self-supervised learning have led to powerful surgical vision encoders capable of spatiotemporal understanding. However, extendin

Capture Timing-Attention of Events in Clinical Time Series

Model ReleasesDGX agent

arXiv:2602.10385v2 Announce Type: replace Abstract: Automatically discovering personalized sequential events from large-scale time-series data is crucial for enabling precision medicine in clinical re

CARI4D: Category Agnostic 4D Reconstruction of Human-Object Interaction

Model ReleasesDGX agent

arXiv:2512.11988v3 Announce Type: replace Abstract: Accurate capture of human-object interaction from ubiquitous sensors like RGB cameras is important for applications in human understanding, gaming,

CaseFacts: A Benchmark for Legal Fact-Checking and Precedent Retrieval

Model ReleasesDGX agent

arXiv:2601.17230v2 Announce Type: replace Abstract: Automated Fact-Checking has largely focused on verifying general knowledge against static corpora, overlooking high-stakes domains like law where tr

CATP: Confidence-Aware Token Pruning for Camouflaged Object Detection

Model ReleasesDGX agent

arXiv:2604.16854v1 Announce Type: new Abstract: Camouflaged Object Detection (COD) aims to segment targets that share extreme textural and structural similarities with their complex environments. Leve

CaTS-Bench: Can Language Models Describe Time Series?

Model ReleasesDGX agent

arXiv:2509.20823v5 Announce Type: replace-cross Abstract: Time series captioning, the task of describing time series in natural language, requires numeric and temporal reasoning, trend interpretation,

CBRS: Cognitive Blood Request System with Bilingual Dataset and Dual-Layer Filtering for Multi-Platform Social Streams

Model ReleasesDGX agent

arXiv:2604.16665v1 Announce Type: new Abstract: Urgent blood donation seeking posts and messages on social media often go unnoticed due to the overwhelming volume of daily communications. Traditional

CDSA-Net:Collaborative Decoupling of Vascular Structure and Background for High-Fidelity Coronary Digital Subtraction Angiography

Model ReleasesDGX agent

arXiv:2604.17208v1 Announce Type: new Abstract: Digital subtraction angiography (DSA) in coronary imaging is fundamentally challenged by physiological motion, forcing reliance on raw angiograms clutte

CFMS: Towards Explainable and Fine-Grained Chinese Multimodal Sarcasm Detection Benchmark

Model ReleasesDGX agent

arXiv:2604.16372v1 Announce Type: new Abstract: Multimodal sarcasm detection has recently garnered significant attention. However, existing benchmarks suffer from coarse-grained annotations and limite

Chain Of Interaction Benchmark (COIN): When Reasoning meets Embodied Interaction

Model ReleasesDGX agent

arXiv:2604.16886v1 Announce Type: new Abstract: Generalist embodied agents must perform interactive, causally-dependent reasoning, continually interacting with the environment, acquiring information,

Channel Attention-Guided Cross-Modal Knowledge Distillation for Referring Image Segmentation

Model ReleasesDGX agent

arXiv:2604.16806v1 Announce Type: new Abstract: Referring image segmentation (RIS) requires accurate segmentation of target regions in images according to language descriptions, which is a cross-modal

Chatting about Upper-Body Expressive Human Pose and Shape Estimation

Model ReleasesDGX agent

arXiv:2604.17959v1 Announce Type: new Abstract: Expressive Human Pose and Shape Estimation (EHPS) plays a crucial role in various AR/VR applications and has witnessed significant progress in recent ye

Claude Opus 4.7 with adaptive thinking via the API... am I missing something or is it not possible any more to force it to think? (Prompt ha…

Model ReleasesDGX agent

Claude Opus 4.7 with adaptive thinking via the API... am I missing something or is it not possible any more to force it to think? (Prompt hacks like 'think step by step' don't count here, I mean the e

ClawEnvKit: Automatic Environment Generation for Claw-Like Agents

Model ReleasesDGX agent

arXiv:2604.18543v1 Announce Type: cross Abstract: Constructing environments for training and evaluating claw-like agents remains a manual, human-intensive process that does not scale. We argue that wh

CodePivot: Bootstrapping Multilingual Transpilation in LLMs via Reinforcement Learning without Parallel Corpora

Model ReleasesDGX agent

arXiv:2604.18027v1 Announce Type: cross Abstract: Transpilation, or code translation, aims to convert source code from one programming language (PL) to another. It is beneficial for many downstream ap

CoDial: Interpretable Task-Oriented Dialogue Systems Through Dialogue Flow Alignment

Model ReleasesDGX agent

arXiv:2506.02264v3 Announce Type: replace Abstract: Building Task-Oriented Dialogue (TOD) systems that generalize across different tasks remains a challenging problem. Data-driven approaches often str

Cognitive Chain-of-Thought (CoCoT): Structured Multimodal Reasoning about Social Situations

Model ReleasesDGX agent

arXiv:2507.20409v2 Announce Type: replace Abstract: Chain-of-Thought (CoT) prompting helps models think step by step. But naive CoT breaks down in visually grounded social tasks, where models must per

CoLLM: A Unified Framework for Co-execution of LLMs Federated Fine-tuning and Inference

Model ReleasesDGX agent

arXiv:2604.16400v1 Announce Type: cross Abstract: As Large Language Models (LLMs) are increasingly adopted in edge intelligence to power domain-specific applications and personalized services, the qua

Combined Hyperbolic and Euclidean Soft Triple Loss Beyond the Single Space Deep Metric Learning

Model ReleasesDGX agent

arXiv:2510.05643v2 Announce Type: replace Abstract: Deep metric learning (DML) aims to learn a neural network mapping data to an embedding space, which can represent semantic similarity between data p

ComPASS: Towards Personalized Agentic Social Support via Tool-Augmented Companionship

Model ReleasesDGX agent

arXiv:2604.18356v1 Announce Type: new Abstract: Developing compassionate interactive systems requires agents to not only understand user emotions but also provide diverse, substantive support. While r

Complex normalizing flows can be information Kahler-Ricci flows

Model ReleasesDGX agent

arXiv:2604.17954v1 Announce Type: cross Abstract: We develop interconnections between the complex normalizing flow for data drawn from Borel probability measures on the twofold realification of the co

Concurrent Criterion Validation of a Validity Screen for LLM Confidence Signals via Selective Prediction

Model ReleasesDGX agent

arXiv:2604.17716v1 Announce Type: new Abstract: The validity screen (Cacioli, 2026d, 2026e) classifies LLM confidence signals as Valid, Indeterminate, or Invalid. We test whether these classifications

Conditional Attribution for Root Cause Analysis in Time-Series Anomaly Detection

Model ReleasesDGX agent

arXiv:2604.17616v1 Announce Type: new Abstract: Root cause analysis (RCA) for time-series anomaly detection is critical for the reliable operation of complex real-world systems. Existing explanation m

Conformal Risk Control under Non-Monotone Losses: Theory and Finite-Sample Guarantees

Model ReleasesDGX agent

arXiv:2604.01502v2 Announce Type: replace-cross Abstract: Conformal risk control (CRC) provides distribution-free guarantees for controlling the expected loss at a user-specified level. Existing theor

ConMeZO: Adaptive Descent-Direction Sampling for Gradient-Free Finetuning of Large Language Models

Model ReleasesDGX agent

arXiv:2511.02757v2 Announce Type: replace Abstract: Zeroth-order or derivative-free optimization (MeZO) is an attractive strategy for finetuning large language models (LLMs) because it eliminates the

Constructive Distortion: Improving MLLMs with Attention-Guided Image Warping

Model ReleasesDGX agent

arXiv:2510.09741v3 Announce Type: replace Abstract: Multimodal large language models (MLLMs) often miss small details and spatial relations in cluttered scenes, leading to errors in fine-grained perce

Continuous Limits of Coupled Flows in Representation Learning

Model ReleasesDGX agent

arXiv:2604.16801v1 Announce Type: new Abstract: While modern representation learning relies heavily on global error signals, decentralized algorithms driven by local interactions offer a fundamental d

Copy First, Translate Later: Interpreting Translation Dynamics in Multilingual Pretraining

Model ReleasesDGX agent

arXiv:2604.17633v1 Announce Type: new Abstract: Large language models exhibit impressive cross-lingual capabilities. However, prior work analyzes this phenomenon through isolated factors and at sparse

CORP: A Multi-Modal Dataset for Campus-Oriented Roadside Perception Tasks

Model ReleasesDGX agent

arXiv:2404.03191v3 Announce Type: replace Abstract: Numerous roadside perception datasets have been introduced to propel advancements in autonomous driving and intelligent transportation systems resea

Counterfactual Modeling with Fine-Tuned LLMs for Health Intervention Design and Sensor Data Augmentation

Model ReleasesDGX agent

arXiv:2601.14590v2 Announce Type: replace Abstract: Counterfactual explanations (CFEs) provide human-centric interpretability by identifying the minimal, actionable changes required to alter a machine

Creating ConLangs to Probe the Metalinguistic Grammatical Knowledge of LLMs

Model ReleasesDGX agent

arXiv:2510.07591v3 Announce Type: replace Abstract: We present a system that uses LLMs as a tool in the development of Constructed Languages -- ConLangs, which we call IASC (Interactive Agentic System

CreditDecoding: Accelerating Parallel Decoding in Diffusion Large Language Models with Trace Credit

Model ReleasesDGX agent

arXiv:2510.06133v2 Announce Type: replace Abstract: Diffusion large language models (dLLMs) generate text through iterative denoising. In commonly adopted parallel decoding schemes, each step confirms

CROC: Evaluating and Training T2I Metrics with Pseudo- and Human-Labeled Contrastive Robustness Checks

Model ReleasesDGX agent

arXiv:2505.11314v2 Announce Type: replace-cross Abstract: The assessment of evaluation metrics (meta-evaluation) is crucial for determining the suitability of existing metrics in text-to-image (T2I) g

Cross-Family Speculative Decoding for Polish Language Models on Apple~Silicon: An Empirical Evaluation of Bielik~11B with UAG-Extended MLX-LM

Model ReleasesDGX agent

arXiv:2604.16368v1 Announce Type: new Abstract: Speculative decoding accelerates LLM inference by using a small draft model to propose k candidate tokens for a target model to verify. While effective

← Previous
1…324325326327328…372
Next →