AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
All
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,340 results
Model Releases

Beyond 'I Don't Know': Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty

DGX agent

arXiv:2604.17293v1 Announce Type: new Abstract: Reliable Large Language Models (LLMs) should abstain when confidence is insufficient. However, prior studies often treat refusal as a generic 'I don't k

model-releasesarxiv-cs-cl
21 Apr 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Beyond Pattern Matching: Seven Cross-Domain Techniques for Prompt Injection Detection

DGX agent

arXiv:2604.18248v1 Announce Type: cross Abstract: Current open-source prompt-injection detectors converge on two architectural choices: regular-expression pattern matching and fine-tuned transformer c

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Beyond Reproduction: A Paired-Task Framework for Assessing LLM Comprehension and Creativity in Literary Translation

DGX agent

arXiv:2604.18169v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for creative tasks such as literary translation. Yet translational creativity remains underexplored a

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Beyond Word Boundaries: A Hebrew Coreference Benchmark and an Evaluation Protocol for Morphologically Complex Text

DGX agent

arXiv:2604.17108v1 Announce Type: new Abstract: Coreference Resolution (CR) is a fundamental NLP task critical for long-form tasks as information extraction, summarization, and many business applicati

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Bi-LoRA: Efficient Sharpness-Aware Minimization for Fine-Tuning Large-Scale Models

DGX agent

arXiv:2508.19564v2 Announce Type: replace Abstract: Fine-tuning large-scale pre-trained models with limited data presents significant challenges for generalization. While Sharpness-Aware Minimization

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Bielik Guard: Efficient Polish Language Safety Classifiers for LLM Content Moderation

DGX agent

arXiv:2602.07954v4 Announce Type: replace Abstract: As Large Language Models (LLMs) become increasingly deployed in Polish language applications, the need for efficient and accurate content safety cla

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

BOOKAGENT: Orchestrating Safety-Aware Visual Narratives via Multi-Agent Cognitive Calibration

DGX agent

arXiv:2604.16541v1 Announce Type: new Abstract: Recent advancements in Large Generative Models (LGMs) have revolutionized multi-modal generation. However, generating illustrated storybooks remains an

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

BOP-ASK: Object-Interaction Reasoning for Vision-Language Models

DGX agent

arXiv:2511.16857v3 Announce Type: replace Abstract: Vision Language Models (VLMs) have achieved impressive performance on spatial reasoning benchmarks, yet these evaluations mask critical weaknesses i

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Brain-Inspired Capture: Evidence-Driven Neuromimetic Perceptual Simulation for Visual Decoding

DGX agent

arXiv:2604.17927v1 Announce Type: new Abstract: Visual decoding of neurophysiological signals is a critical challenge for brain-computer interfaces (BCIs) and computational neuroscience. However, curr

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

BridgeEQA: Virtual Embodied Agents for Real Bridge Inspections

DGX agent

arXiv:2511.12676v2 Announce Type: replace Abstract: Deploying embodied agents that can answer questions about their surroundings in realistic real-world settings remains difficult, partly due to the s

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Bridging the Reasoning Gap in Vietnamese with Small Language Models via Test-Time Scaling

DGX agent

arXiv:2604.17794v1 Announce Type: new Abstract: The democratization of ubiquitous AI hinges on deploying sophisticated reasoning capabilities on resource-constrained devices. However, Small Language M

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Budget-Aware Anytime Reasoning with LLM-Synthesized Preference Data

DGX agent

arXiv:2601.11038v2 Announce Type: replace Abstract: We study the reasoning behavior of large language models (LLMs) under limited computation budgets. In such settings, producing useful partial soluti

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

CAM3DNet: Comprehensively mining the multi-scale features for 3D Object Detection with Multi-View Cameras

DGX agent

arXiv:2604.17024v1 Announce Type: new Abstract: Query-based 3D object detection methods using multi-view images often struggle to efficiently leverage dynamic multi-scale information, e.g., the relati

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Camo-M3FD: A New Benchmark Dataset for Cross-Spectral Camouflaged Pedestrian Detection

DGX agent

arXiv:2604.16582v1 Announce Type: new Abstract: Pedestrian detection is fundamental to autonomous driving, robotics, and surveillance. Despite progress in deep learning, reliable identification remain

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Can Large Language Models Understand Context?

DGX agent

Understanding context is key to understanding human language, an ability which Large Language Models (LLMs) have been increasingly seen to demonstrate to an impressive extent. However, though the eval

model-releasesapple-ml-research
21 Apr 2026
Model Releases

Can LLM-Generated Text Empower Surgical Vision-Language Pre-training?

DGX agent

arXiv:2604.18134v1 Announce Type: new Abstract: Recent advancements in self-supervised learning have led to powerful surgical vision encoders capable of spatiotemporal understanding. However, extendin

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Capture Timing-Attention of Events in Clinical Time Series

DGX agent

arXiv:2602.10385v2 Announce Type: replace Abstract: Automatically discovering personalized sequential events from large-scale time-series data is crucial for enabling precision medicine in clinical re

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

CARI4D: Category Agnostic 4D Reconstruction of Human-Object Interaction

DGX agent

arXiv:2512.11988v3 Announce Type: replace Abstract: Accurate capture of human-object interaction from ubiquitous sensors like RGB cameras is important for applications in human understanding, gaming,

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

CaseFacts: A Benchmark for Legal Fact-Checking and Precedent Retrieval

DGX agent

arXiv:2601.17230v2 Announce Type: replace Abstract: Automated Fact-Checking has largely focused on verifying general knowledge against static corpora, overlooking high-stakes domains like law where tr

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

CATP: Confidence-Aware Token Pruning for Camouflaged Object Detection

DGX agent

arXiv:2604.16854v1 Announce Type: new Abstract: Camouflaged Object Detection (COD) aims to segment targets that share extreme textural and structural similarities with their complex environments. Leve

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

CaTS-Bench: Can Language Models Describe Time Series?

DGX agent

arXiv:2509.20823v5 Announce Type: replace-cross Abstract: Time series captioning, the task of describing time series in natural language, requires numeric and temporal reasoning, trend interpretation,

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

CBRS: Cognitive Blood Request System with Bilingual Dataset and Dual-Layer Filtering for Multi-Platform Social Streams

DGX agent

arXiv:2604.16665v1 Announce Type: new Abstract: Urgent blood donation seeking posts and messages on social media often go unnoticed due to the overwhelming volume of daily communications. Traditional

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

CDSA-Net:Collaborative Decoupling of Vascular Structure and Background for High-Fidelity Coronary Digital Subtraction Angiography

DGX agent

arXiv:2604.17208v1 Announce Type: new Abstract: Digital subtraction angiography (DSA) in coronary imaging is fundamentally challenged by physiological motion, forcing reliance on raw angiograms clutte

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

CFMS: Towards Explainable and Fine-Grained Chinese Multimodal Sarcasm Detection Benchmark

DGX agent

arXiv:2604.16372v1 Announce Type: new Abstract: Multimodal sarcasm detection has recently garnered significant attention. However, existing benchmarks suffer from coarse-grained annotations and limite

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Chain Of Interaction Benchmark (COIN): When Reasoning meets Embodied Interaction

DGX agent

arXiv:2604.16886v1 Announce Type: new Abstract: Generalist embodied agents must perform interactive, causally-dependent reasoning, continually interacting with the environment, acquiring information,

model-releasesarxiv-cs-ro
21 Apr 2026
Model Releases

Channel Attention-Guided Cross-Modal Knowledge Distillation for Referring Image Segmentation

DGX agent

arXiv:2604.16806v1 Announce Type: new Abstract: Referring image segmentation (RIS) requires accurate segmentation of target regions in images according to language descriptions, which is a cross-modal

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Chatting about Upper-Body Expressive Human Pose and Shape Estimation

DGX agent

arXiv:2604.17959v1 Announce Type: new Abstract: Expressive Human Pose and Shape Estimation (EHPS) plays a crucial role in various AR/VR applications and has witnessed significant progress in recent ye

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Claude Opus 4.7 with adaptive thinking via the API... am I missing something or is it not possible any more to force it to think? (Prompt ha…

DGX agent

Claude Opus 4.7 with adaptive thinking via the API... am I missing something or is it not possible any more to force it to think? (Prompt hacks like 'think step by step' don't count here, I mean the e

model-releasessimon-willison--x
21 Apr 2026
Model Releases

ClawEnvKit: Automatic Environment Generation for Claw-Like Agents

DGX agent

arXiv:2604.18543v1 Announce Type: cross Abstract: Constructing environments for training and evaluating claw-like agents remains a manual, human-intensive process that does not scale. We argue that wh

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

CodePivot: Bootstrapping Multilingual Transpilation in LLMs via Reinforcement Learning without Parallel Corpora

DGX agent

arXiv:2604.18027v1 Announce Type: cross Abstract: Transpilation, or code translation, aims to convert source code from one programming language (PL) to another. It is beneficial for many downstream ap

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

CoDial: Interpretable Task-Oriented Dialogue Systems Through Dialogue Flow Alignment

DGX agent

arXiv:2506.02264v3 Announce Type: replace Abstract: Building Task-Oriented Dialogue (TOD) systems that generalize across different tasks remains a challenging problem. Data-driven approaches often str

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Cognitive Chain-of-Thought (CoCoT): Structured Multimodal Reasoning about Social Situations

DGX agent

arXiv:2507.20409v2 Announce Type: replace Abstract: Chain-of-Thought (CoT) prompting helps models think step by step. But naive CoT breaks down in visually grounded social tasks, where models must per

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

CoLLM: A Unified Framework for Co-execution of LLMs Federated Fine-tuning and Inference

DGX agent

arXiv:2604.16400v1 Announce Type: cross Abstract: As Large Language Models (LLMs) are increasingly adopted in edge intelligence to power domain-specific applications and personalized services, the qua

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Combined Hyperbolic and Euclidean Soft Triple Loss Beyond the Single Space Deep Metric Learning

DGX agent

arXiv:2510.05643v2 Announce Type: replace Abstract: Deep metric learning (DML) aims to learn a neural network mapping data to an embedding space, which can represent semantic similarity between data p

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

ComPASS: Towards Personalized Agentic Social Support via Tool-Augmented Companionship

DGX agent

arXiv:2604.18356v1 Announce Type: new Abstract: Developing compassionate interactive systems requires agents to not only understand user emotions but also provide diverse, substantive support. While r

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Complex normalizing flows can be information Kahler-Ricci flows

DGX agent

arXiv:2604.17954v1 Announce Type: cross Abstract: We develop interconnections between the complex normalizing flow for data drawn from Borel probability measures on the twofold realification of the co

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Concurrent Criterion Validation of a Validity Screen for LLM Confidence Signals via Selective Prediction

DGX agent

arXiv:2604.17716v1 Announce Type: new Abstract: The validity screen (Cacioli, 2026d, 2026e) classifies LLM confidence signals as Valid, Indeterminate, or Invalid. We test whether these classifications

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Conditional Attribution for Root Cause Analysis in Time-Series Anomaly Detection

DGX agent

arXiv:2604.17616v1 Announce Type: new Abstract: Root cause analysis (RCA) for time-series anomaly detection is critical for the reliable operation of complex real-world systems. Existing explanation m

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Conformal Risk Control under Non-Monotone Losses: Theory and Finite-Sample Guarantees

DGX agent

arXiv:2604.01502v2 Announce Type: replace-cross Abstract: Conformal risk control (CRC) provides distribution-free guarantees for controlling the expected loss at a user-specified level. Existing theor

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

ConMeZO: Adaptive Descent-Direction Sampling for Gradient-Free Finetuning of Large Language Models

DGX agent

arXiv:2511.02757v2 Announce Type: replace Abstract: Zeroth-order or derivative-free optimization (MeZO) is an attractive strategy for finetuning large language models (LLMs) because it eliminates the

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Constructive Distortion: Improving MLLMs with Attention-Guided Image Warping

DGX agent

arXiv:2510.09741v3 Announce Type: replace Abstract: Multimodal large language models (MLLMs) often miss small details and spatial relations in cluttered scenes, leading to errors in fine-grained perce

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Continuous Limits of Coupled Flows in Representation Learning

DGX agent

arXiv:2604.16801v1 Announce Type: new Abstract: While modern representation learning relies heavily on global error signals, decentralized algorithms driven by local interactions offer a fundamental d

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Copy First, Translate Later: Interpreting Translation Dynamics in Multilingual Pretraining

DGX agent

arXiv:2604.17633v1 Announce Type: new Abstract: Large language models exhibit impressive cross-lingual capabilities. However, prior work analyzes this phenomenon through isolated factors and at sparse

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

CORP: A Multi-Modal Dataset for Campus-Oriented Roadside Perception Tasks

DGX agent

arXiv:2404.03191v3 Announce Type: replace Abstract: Numerous roadside perception datasets have been introduced to propel advancements in autonomous driving and intelligent transportation systems resea

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Counterfactual Modeling with Fine-Tuned LLMs for Health Intervention Design and Sensor Data Augmentation

DGX agent

arXiv:2601.14590v2 Announce Type: replace Abstract: Counterfactual explanations (CFEs) provide human-centric interpretability by identifying the minimal, actionable changes required to alter a machine

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Creating ConLangs to Probe the Metalinguistic Grammatical Knowledge of LLMs

DGX agent

arXiv:2510.07591v3 Announce Type: replace Abstract: We present a system that uses LLMs as a tool in the development of Constructed Languages -- ConLangs, which we call IASC (Interactive Agentic System

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

CreditDecoding: Accelerating Parallel Decoding in Diffusion Large Language Models with Trace Credit

DGX agent

arXiv:2510.06133v2 Announce Type: replace Abstract: Diffusion large language models (dLLMs) generate text through iterative denoising. In commonly adopted parallel decoding schemes, each step confirms

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

CROC: Evaluating and Training T2I Metrics with Pseudo- and Human-Labeled Contrastive Robustness Checks

DGX agent

arXiv:2505.11314v2 Announce Type: replace-cross Abstract: The assessment of evaluation metrics (meta-evaluation) is crucial for determining the suitability of existing metrics in text-to-image (T2I) g

model-releasesarxiv-cs-cl
21 Apr 2026
← Previous
1…407408409410411…466
Next →