AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,280 results
14 Apr 2026

Conflicts Make Large Reasoning Models Vulnerable to Attacks

Model ReleasesDGX agent

arXiv:2604.09750v1 Announce Type: cross Abstract: Large Reasoning Models (LRMs) have achieved remarkable performance across diverse domains, yet their decision-making under conflicting objectives rema

Consistency of AI-Generated Exercise Prescriptions: A Repeated Generation Study Using a Large Language Model

Model ReleasesDGX agent

arXiv:2604.11287v1 Announce Type: new Abstract: Background: Large language models (LLMs) have been explored as tools for generating personalized exercise prescriptions, yet the consistency of outputs

Context-Aware Semantic Segmentation via Stage-Wise Attention

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2601.11310v2 Announce Type: replace Abstract: Semantic ultra-high-resolution (UHR) image segmentation is essential in remote sensing applications such as aerial mapping and environmental monitor

CounterBench: Evaluating and Improving Counterfactual Reasoning in Large Language Models

Model ReleasesDGX agent

arXiv:2502.11008v2 Announce Type: replace Abstract: Counterfactual reasoning is widely recognized as one of the most challenging and intricate aspects of causality in artificial intelligence. In this

Counting to Four is still a Chore for VLMs

Model ReleasesDGX agent

arXiv:2604.10039v1 Announce Type: new Abstract: Vision--language models (VLMs) have achieved impressive performance on complex multimodal reasoning tasks, yet they still fail on simple grounding skill

CPAM: Context-Preserving Adaptive Manipulation for Zero-Shot Real Image Editing

Model ReleasesDGX agent

arXiv:2506.18438v2 Announce Type: replace Abstract: Editing natural images using textual descriptions in text-to-image diffusion models remains a significant challenge, particularly in achieving consi

CricBench: A Multilingual Benchmark for Evaluating LLMs in Cricket Analytics

Model ReleasesDGX agent

arXiv:2512.21877v3 Announce Type: replace-cross Abstract: Cricket is the second most popular sport worldwide, with billions of fans seeking advanced statistical insights unavailable through standard w

Cross-Validated Cross-Channel Self-Attention and Denoising for Automatic Modulation Classification

Model ReleasesDGX agent

arXiv:2604.10054v1 Announce Type: new Abstract: This study addresses a key limitation in deep learning Automatic Modulation Classification (AMC) models, which perform well at high signal-to-noise rati

Cybersecurity Looks Like Proof of Work Now

Model ReleasesDGX agent

Cybersecurity Looks Like Proof of Work Now The UK's AI Safety Institute recently published Our evaluation of Claude Mythos Preview’s cyber capabilities, their own independent analysis of Claude Mythos

Data-Efficient Semantic Segmentation of 3D Point Clouds via Open-Vocabulary Image Segmentation-based Pseudo-Labeling

Model ReleasesDGX agent

arXiv:2604.11007v1 Announce Type: new Abstract: Semantic segmentation of 3D point cloud scenes is a crucial task for various applications. In real-world scenarios, training segmentation models often f

datasette PR #2689: Replace token-based CSRF with Sec-Fetch-Site header protection

Model ReleasesDGX agent

datasette PR #2689: Replace token-based CSRF with Sec-Fetch-Site header protection Datasette has long protected against CSRF attacks using CSRF tokens, implemented using my asgi-csrf Python library. T

DDO-RM for LLM Preference Optimization: A Minimal Held-Out Benchmark against DPO

Model ReleasesDGX agent

arXiv:2604.11119v1 Announce Type: cross Abstract: This paper reorganizes the current manuscript around the DPO versus DDO-RM preference-optimization project and focuses on two parts: the algorithmic v

Dead Cognitions: A Census of Misattributed Insights

Model ReleasesDGX agent

arXiv:2604.10288v1 Announce Type: new Abstract: This essay identifies a failure mode of AI chat systems that we term attribution laundering: the model performs substantive cognitive work and then rhet

DecepGPT: Schema-Driven Deception Detection with Multicultural Datasets and Robust Multimodal Learning

Model ReleasesDGX agent

arXiv:2603.23916v2 Announce Type: replace-cross Abstract: Multimodal deception detection aims to identify deceptive behavior by analyzing audiovisual cues for forensics and security. In these high-sta

Decompose, Mix, Adapt: A Unified Framework for Parameter-Efficient Neural Network Recombination and Compression

Model ReleasesDGX agent

arXiv:2603.27383v2 Announce Type: replace Abstract: Parameter Recombination (PR) methods aim to efficiently compose the weights of a neural network for applications like Parameter-Efficient FineTuning

Decomposing and Reducing Hidden Measurement Error in LLM Evaluation Pipelines

Model ReleasesDGX agent

arXiv:2604.11581v1 Announce Type: new Abstract: LLM evaluations drive which models get deployed, which safety standards get adopted, and which research conclusions get published. Yet these scores carr

🚀 deepagents 0.5 release 👉 Async subagents - kick off background tasks on any Agent Protocol backed server while you continue to interact …

Model ReleasesDGX agent

🚀 deepagents 0.5 release 👉 Async subagents - kick off background tasks on any Agent Protocol backed server while you continue to interact with the main agent. Start multiple background tasks in parall

DeepReviewer 2.0: A Traceable Agentic System for Auditable Scientific Peer Review

Model ReleasesDGX agent

arXiv:2604.09590v1 Announce Type: new Abstract: Automated peer review is often framed as generating fluent critique, yet reviewers and area chairs need judgments they can audit: where a concern applie

Degradation-Consistent Paired Training for Robust AI-Generated Image Detection

Model ReleasesDGX agent

arXiv:2604.10102v1 Announce Type: cross Abstract: AI-generated image detectors suffer significant performance degradation under real-world image corruptions such as JPEG compression, Gaussian blur, an

Delta Rectified Flow Sampling for Text-to-Image Editing

Model ReleasesDGX agent

arXiv:2509.05342v3 Announce Type: replace Abstract: We propose Delta Rectified Flow Sampling (DRFS), a novel inversion-free, path-aware editing framework within rectified flow models for text-to-image

DERM-3R: A Resource-Efficient Multimodal Agents Framework for Dermatologic Diagnosis and Treatment in Real-World Clinical Settings

Model ReleasesDGX agent

arXiv:2604.09596v1 Announce Type: new Abstract: Dermatologic diseases impose a large and growing global burden, affecting billions and substantially reducing quality of life. While modern therapies ca

Design Principles for Sequence Models via Coefficient Dynamics

Model ReleasesDGX agent

arXiv:2510.09389v2 Announce Type: replace-cross Abstract: Deep sequence models, ranging from Transformers and State Space Models (SSMs) to more recent approaches such as gated linear RNNs, fundamental

Detecting Corporate AI-Washing via Cross-Modal Semantic Inconsistency Learning

Model ReleasesDGX agent

arXiv:2604.09644v1 Announce Type: cross Abstract: Corporate AI-washing-the strategic misrepresentation of AI capabilities via exaggerated or fabricated cross-channel disclosures-has emerged as a syste

Detecting critical treatment effect bias in small subgroups

Model ReleasesDGX agent

arXiv:2404.18905v3 Announce Type: replace-cross Abstract: Randomized trials are considered the gold standard for making informed decisions in medicine, yet they often lack generalizability to the pati

Detecting Safety Violations Across Many Agent Traces

Model ReleasesDGX agent

arXiv:2604.11806v1 Announce Type: new Abstract: To identify safety violations, auditors often search over large sets of agent traces. This search is difficult because failures are often rare, complex,

Development and evaluation of CADe systems in low-prevalence setting: The RARE25 challenge for early detection of Barrett's neoplasia

Model ReleasesDGX agent

arXiv:2604.11171v1 Announce Type: new Abstract: Computer-aided detection (CADe) of early neoplasia in Barrett's esophagus is a low-prevalence surveillance problem in which clinically relevant findings

Differentially Private Verification of Distribution Properties

Model ReleasesDGX agent

arXiv:2604.10819v1 Announce Type: cross Abstract: A recent line of work initiated by Chiesa and Gur and further developed by Herman and Rothblum investigates the sample and communication complexity of

DiningBench: A Hierarchical Multi-view Benchmark for Perception and Reasoning in the Dietary Domain

Model ReleasesDGX agent

arXiv:2604.10425v1 Announce Type: new Abstract: Recent advancements in Vision-Language Models (VLMs) have revolutionized general visual understanding. However, their application in the food domain rem

Disambiguation-Centric Finetuning Makes Enterprise Tool-Calling LLMs More Realistic and Less Risky

Model ReleasesDGX agent

arXiv:2507.03336v4 Announce Type: replace Abstract: Large language models (LLMs) are increasingly tasked with invoking enterprise APIs, yet they routinely falter when near-duplicate tools vie for the

DiSPA: Differential Substructure-Pathway Attention for Drug Response Prediction

Model ReleasesDGX agent

arXiv:2601.14346v2 Announce Type: replace-cross Abstract: Accurate prediction of drug response in precision medicine requires models that capture how specific chemical substructures interact with cell

Do Agent Rules Shape or Distort? Guardrails Beat Guidance in Coding Agents

Model ReleasesDGX agent

arXiv:2604.11088v1 Announce Type: new Abstract: Developers increasingly guide AI coding agents through natural language instruction files (e.g., CLAUDE.md, .cursorrules), yet no controlled study has m

Do LLMs Build Spatial World Models? Evidence from Grid-World Maze Tasks

Model ReleasesDGX agent

arXiv:2604.10690v1 Announce Type: new Abstract: Foundation models have shown remarkable performance across diverse tasks, yet their ability to construct internal spatial world models for reasoning and

Do Thought Streams Matter? Evaluating Reasoning in Gemini Vision-Language Models for Video Scene Understanding

Model ReleasesDGX agent

arXiv:2604.11177v1 Announce Type: new Abstract: We benchmark how internal reasoning traces, which we call thought streams, affect video scene understanding in vision-language models. Using four config

Do We Still Need GraphRAG? Benchmarking RAG and GraphRAG for Agentic Search Systems

Model ReleasesDGX agent

arXiv:2604.09666v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) and its graph-based extensions (GraphRAG) are effective paradigms for improving large language model (LLM) reason

Doc-PP: Document Policy Preservation Benchmark for Large Vision-Language Models

Model ReleasesDGX agent

arXiv:2601.03926v2 Announce Type: replace Abstract: The deployment of Large Vision-Language Models (LVLMs) for real-world document question answering is often constrained by dynamic, user-defined poli

DocRevive: A Unified Pipeline for Document Text Restoration

Model ReleasesDGX agent

arXiv:2604.10077v1 Announce Type: new Abstract: In Document Understanding, the challenge of reconstructing damaged, occluded, or incomplete text remains a critical yet unexplored problem. Subsequent d

Domain-Aware Hybrid Quantum Learning via Correlation-Guided Circuit Design for Crime Pattern Analytics

Model ReleasesDGX agent

arXiv:2604.07389v2 Announce Type: replace Abstract: Crime pattern analysis is critical for law enforcement and predictive policing, yet the surge in criminal activities from rapid urbanization creates

DoReMi: Bridging 3D Domains via Topology-Aware Domain-Representation Mixture of Experts

Model ReleasesDGX agent

arXiv:2511.11232v2 Announce Type: replace Abstract: Constructing a unified 3D scene understanding model has long been hindered by the significant topological discrepancies across different sensor moda

DPNet: Doppler LiDAR Motion Planning for Highly-Dynamic Environments

Model ReleasesDGX agent

arXiv:2512.00375v2 Announce Type: replace Abstract: Existing motion planning methods often struggle with rapid-motion obstacles due to an insufficient understanding of environmental changes. To addres

E2E-REME: Towards End-to-End Microservices Auto-Remediation via Experience-Simulation Reinforcement Fine-Tuning

Model ReleasesDGX agent

arXiv:2604.11094v1 Announce Type: cross Abstract: Contemporary microservice systems continue to grow in scale and complexity, leading to increasingly frequent and costly failures. While recent LLM-bas

EagleVision: A Multi-Task Benchmark for Cross-Domain Perception in High-Speed Autonomous Racing

Model ReleasesDGX agent

arXiv:2604.11400v1 Announce Type: cross Abstract: High-speed autonomous racing presents extreme perception challenges, including large relative velocities and substantial domain shifts from convention

EdgeCIM: A Hardware-Software Co-Design for CIM-Based Acceleration of Small Language Models

Model ReleasesDGX agent

arXiv:2604.11512v1 Announce Type: cross Abstract: The growing demand for deploying Small Language Models (SLMs) on edge devices, including laptops, smartphones, and embedded platforms, has exposed fun

EdgeDAM: Real-time Object Tracking for Mobile Devices

Model ReleasesDGX agent

arXiv:2603.05463v2 Announce Type: replace Abstract: Single-object tracking (SOT) on edge devices is a critical computer vision task, requiring accurate and continuous target localization across video

Edu-MMBias: A Three-Tier Multimodal Benchmark for Auditing Social Bias in Vision-Language Models under Educational Contexts

Model ReleasesDGX agent

arXiv:2604.10200v1 Announce Type: new Abstract: As Vision-Language Models (VLMs) become integral to educational decision-making, ensuring their fairness is paramount. However, current text-centric eva

EduIllustrate: Towards Scalable Automated Generation Of Multimodal Educational Content

Model ReleasesDGX agent

arXiv:2604.05005v2 Announce Type: replace-cross Abstract: Large language models are increasingly used as educational assistants, yet evaluation of their educational capabilities remains concentrated o

Eevee: Towards Close-up High-resolution Video-based Virtual Try-on

Model ReleasesDGX agent

arXiv:2511.18957v2 Announce Type: replace Abstract: Video virtual try-on technology provides a cost-effective solution for creating marketing videos in fashion e-commerce. However, its practical adopt

EgoFun3D: Modeling Interactive Objects from Egocentric Videos using Function Templates

Model ReleasesDGX agent

arXiv:2604.11038v1 Announce Type: new Abstract: We present EgoFun3D, a coordinated task formulation, dataset, and benchmark for modeling interactive 3D objects from egocentric videos. Interactive obje

Eliciting Medical Reasoning with Knowledge-enhanced Data Synthesis: A Semi-Supervised Reinforcement Learning Approach

Model ReleasesDGX agent

arXiv:2604.11547v1 Announce Type: cross Abstract: While large language models hold promise for complex medical applications, their development is hindered by the scarcity of high-quality reasoning dat

EmbodiedGovBench: A Benchmark for Governance, Recovery, and Upgrade Safety in Embodied Agent Systems

Model ReleasesDGX agent

arXiv:2604.11174v1 Announce Type: cross Abstract: Recent progress in embodied AI has produced a growing ecosystem of robot policies, foundation models, and modular runtimes. However, current evaluatio

End-to-end Automated Deep Neural Network Optimization for PPG-based Blood Pressure Estimation on Wearables

Model ReleasesDGX agent

arXiv:2604.10117v1 Announce Type: new Abstract: Photoplethysmography (PPG)-based blood pressure (BP) estimation is a challenging task, particularly on resource-constrained wearable devices. However, f

Enhanced-FQL(lambda), an Efficient and Interpretable RL with novel Fuzzy Eligibility Traces and Segmented Experience Replay

Model ReleasesDGX agent

arXiv:2601.04392v2 Announce Type: replace-cross Abstract: This paper introduces a fuzzy reinforcement learning framework, Enhanced-FQL(lambda), that integrates novel Fuzzified Eligibility Traces (FET)

Enhancing Multimodal Large Language Models for Ancient Chinese Character Evolution Analysis via Glyph-Driven Fine-Tuning

Model ReleasesDGX agent

arXiv:2604.11299v1 Announce Type: cross Abstract: In recent years, rapid advances in Multimodal Large Language Models (MLLMs) have increasingly stimulated research on ancient Chinese scripts. As the e

Environmental Footprint of GenAI Research: Insights from the Moshi Foundation Model

Model ReleasesDGX agent

arXiv:2604.11154v1 Announce Type: new Abstract: New multi-modal large language models (MLLMs) are continuously being trained and deployed, following rapid development cycles. This generative AI frenzy

ERNIE Image released

Model ReleasesDGX agent

ERNIE Image is an open-source text-to-image generation model developed by Baidu, built on a single-stream Diffusion Transformer (DiT) paired with a lightweight Prompt Enhancer that expands brief user

Escaping the Context Bottleneck: Active Context Curation for LLM Agents via Reinforcement Learning

Model ReleasesDGX agent

arXiv:2604.11462v1 Announce Type: new Abstract: Large Language Models (LLMs) struggle with long-horizon tasks due to the 'context bottleneck' and the 'lost-in-the-middle' phenomenon, where accumulated

Evaluating Memory Capability in Continuous Lifelog Scenario

Model ReleasesDGX agent

arXiv:2604.11182v1 Announce Type: new Abstract: Nowadays, wearable devices can continuously lifelog ambient conversations, creating substantial opportunities for memory systems. However, existing benc

Evaluating Reliability Gaps in Large Language Model Safety via Repeated Prompt Sampling

Model ReleasesDGX agent

arXiv:2604.09606v1 Announce Type: new Abstract: Traditional benchmarks for large language models (LLMs), such as HELM and AIR-BENCH, primarily assess safety risk through breadth-oriented evaluation ac

Evaluating Scene-based In-Situ Item Labeling for Immersive Conversational Recommendation

Model ReleasesDGX agent

arXiv:2604.09698v1 Announce Type: cross Abstract: The growing ubiquity of Extended Reality (XR) is driving Conversational Recommendation Systems (CRS) toward visually immersive experiences. We formali

Evaluating Small Open LLMs for Medical Question Answering: A Practical Framework

Model ReleasesDGX agent

arXiv:2604.10535v1 Announce Type: cross Abstract: Incorporating large language models (LLMs) in medical question answering demands more than high average accuracy: a model that returns substantively d

EviRCOD: Evidence-Guided Probabilistic Decoding for Referring Camouflaged Object Detection

Model ReleasesDGX agent

arXiv:2604.10894v1 Announce Type: new Abstract: Referring Camouflaged Object Detection (Ref-COD) focuses on segmenting specific camouflaged targets in a query image using category-aligned references.

← Previous
1…350351352353354…372
Next →