AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,272 results
Model Releases

Pando: Do Interpretability Methods Work When Models Won't Explain Themselves?

DGX agent

arXiv:2604.11061v1 Announce Type: cross Abstract: Mechanistic interpretability is often motivated for alignment auditing, where a model's verbal explanations can be absent, incomplete, or misleading.

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Panoptic Pairwise Distortion Graph

DGX agent

arXiv:2604.11004v1 Announce Type: cross Abstract: In this work, we introduce a new perspective on comparative image assessment by representing an image pair as a structured composition of its regions.

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

PaperScope: A Multi-Modal Multi-Document Benchmark for Agentic Deep Research Across Massive Scientific Papers

DGX agent

arXiv:2604.11307v1 Announce Type: new Abstract: Leveraging Multi-modal Large Language Models (MLLMs) to accelerate frontier scientific research is promising, yet how to rigorously evaluate such system

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Para-B&B: Load-Balanced Deterministic Parallelization of Solving MIP

DGX agent

arXiv:2604.09556v1 Announce Type: cross Abstract: Mixed-integer programming (MIP) extends linear programming by incorporating both continuous and integer decision variables, making it widely used in p

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Parameter Efficient Fine-tuning for Domain-specific Gastrointestinal Disease Recognition

DGX agent

arXiv:2604.10451v1 Announce Type: new Abstract: Despite recent advancements in the field of medical image analysis with the use of pretrained foundation models, the issue of distribution shifts betwee

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

ParseBench is the most comprehensive OCR benchmark for real-world enterprise documents: financial filings, contracts, insurance documents, a…

DGX agent

ParseBench is the most comprehensive OCR benchmark for real-world enterprise documents: financial filings, contracts, insurance documents, and more. We evaluate across 5 dimensions that are present am

model-releasesjerry-liu--x
14 Apr 2026
Model Releases

PepBenchmark: A Standardized Benchmark for Peptide Machine Learning

DGX agent

arXiv:2604.10531v1 Announce Type: cross Abstract: Peptide therapeutics are widely regarded as the 'third generation' of drugs, yet progress in peptide Machine Learning (ML) are hindered by the absence

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Persona Non Grata: Single-Method Safety Evaluation Is Incomplete for Persona-Imbued LLMs

DGX agent

arXiv:2604.11120v1 Announce Type: new Abstract: Personality imbuing customizes LLM behavior, but safety evaluations almost always study prompt-based personas alone. We show this is incomplete: prompti

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

PhyMix: Towards Physically Consistent Single-Image 3D Indoor Scene Generation with Implicit--Explicit Optimization

DGX agent

arXiv:2604.10125v1 Announce Type: new Abstract: Existing single-image 3D indoor scene generators often produce results that look visually plausible but fail to obey real-world physics, limiting their

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Physics-informed AI Accelerated Retention Analysis of Ferroelectric Vertical NAND: From Day-Scale TCAD to Second-Scale Surrogate Model

DGX agent

arXiv:2603.06881v2 Announce Type: replace-cross Abstract: Ferroelectric field-effect transistors (FeFET)-based vertical NAND (Fe-VNAND) has emerged as a promising candidate to overcome z-scaling limit

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Pioneer Agent: Continual Improvement of Small Language Models in Production

DGX agent

arXiv:2604.09791v1 Announce Type: new Abstract: Small language models are attractive for production deployment due to their low cost, fast inference, and ease of specialization. However, adapting them

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Playing Along: Learning a Double-Agent Defender for Belief Steering via Theory of Mind

DGX agent

arXiv:2604.11666v1 Announce Type: cross Abstract: As large language models (LLMs) become the engine behind conversational systems, their ability to reason about the intentions and states of their dial

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Please Make it Sound like Human: Encoder-Decoder vs. Decoder-Only Transformers for AI-to-Human Text Style Transfer

DGX agent

arXiv:2604.11687v1 Announce Type: new Abstract: AI-generated text has become common in academic and professional writing, prompting research into detection methods. Less studied is the reverse: system

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Point2Pose: Occlusion-Recovering 6D Pose Tracking and 3D Reconstruction for Multiple Unknown Objects Via 2D Point Trackers

DGX agent

arXiv:2604.10415v1 Announce Type: new Abstract: We present Point2Pose, a model-free method for causal 6D pose tracking of multiple rigid objects from monocular RGB-D video. Initialized only from spars

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

PokeRL: Reinforcement Learning for Pokemon Red

DGX agent

arXiv:2604.10812v1 Announce Type: new Abstract: Pokemon Red is a long-horizon JRPG with sparse rewards, partial observability, and quirky control mechanics that make it a challenging benchmark for rei

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

Policy-Guided Threat Hunting: An LLM enabled Framework with Splunk SOC Triage

DGX agent

arXiv:2603.23966v3 Announce Type: replace-cross Abstract: With frequently evolving Advanced Persistent Threats (APTs) in cyberspace, traditional security solutions approaches have become inadequate fo

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Polyglot Teachers: Evaluating Language Models for Multilingual Synthetic Data Generation

DGX agent

arXiv:2604.11290v1 Announce Type: new Abstract: Synthesizing supervised finetuning (SFT) data from language models (LMs) to teach smaller models multilingual tasks has become increasingly common. Howe

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Powerful Training-Free Membership Inference Against Autoregressive Language Models

DGX agent

arXiv:2601.12104v2 Announce Type: replace-cross Abstract: Fine-tuned language models pose significant privacy risks, as they may memorize and expose sensitive information from their training data. Mem

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Precision Synthesis of Multi-Tracer PET via VLM-Modulated Rectified Flow for Stratifying Mild Cognitive Impairment

DGX agent

arXiv:2604.11176v1 Announce Type: new Abstract: The biological definition of Alzheimer's disease (AD) relies on multi-modal neuroimaging, yet the clinical utility of positron emission tomography (PET)

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

ProGAL-VLA: Grounded Alignment through Prospective Reasoning in Vision-Language-Action Models

DGX agent

arXiv:2604.09824v1 Announce Type: cross Abstract: Vision language action (VLA) models enable generalist robotic agents but often exhibit language ignorance, relying on visual shortcuts and remaining i

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Progressive Multimodal Interaction Network for Reliable Quantification of Fish Feeding Intensity in Aquaculture

DGX agent

arXiv:2506.14170v3 Announce Type: replace-cross Abstract: Accurate quantification of fish feeding intensity is crucial for precision feeding in aquaculture, as it directly affects feed utilization and

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

PSF-Med: Measuring and Explaining Paraphrase Sensitivity in Medical Vision Language Models

DGX agent

arXiv:2602.21428v2 Announce Type: replace Abstract: Medical Vision Language Models (VLMs) can change their answers when clinicians rephrase the same question, a failure mode that threatens deployment

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Quantitative Introspection in Language Models: Tracking Emotive States Across Conversation

DGX agent

arXiv:2603.18893v2 Announce Type: replace Abstract: Tracking the internal states of large language models across conversations is important for safety, interpretability, and model welfare, yet current

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Quantization Dominates Rank Reduction for KV-Cache Compression

DGX agent

arXiv:2604.11501v1 Announce Type: cross Abstract: We compare two strategies for compressing the KV cache in transformer inference: rank reduction (discard dimensions) and quantization (keep all dimens

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Radiology Report Generation for Low-Quality X-Ray Images

DGX agent

arXiv:2604.10188v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have significantly advanced automated Radiology Report Generation (RRG). However, existing methods implicitly assume high-

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

RationalRewards: Reasoning Rewards Scale Visual Generation Both Training and Test Time

DGX agent

arXiv:2604.11626v1 Announce Type: new Abstract: Most reward models for visual generation reduce rich human judgments to a single unexplained score, discarding the reasoning that underlies preference.

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

RCBSF: A Multi-Agent Framework for Automated Contract Revision via Stackelberg Game

DGX agent

arXiv:2604.10740v1 Announce Type: new Abstract: Despite the widespread adoption of Large Language Models (LLMs) in Legal AI, their utility for automated contract revision remains impeded by hallucinat

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Reasoning as Gradient: Scaling MLE Agents Beyond Tree Search

DGX agent

arXiv:2603.01692v3 Announce Type: replace-cross Abstract: LLM-based agents for machine learning engineering (MLE) predominantly rely on tree search, a form of gradient-free optimization that uses scal

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

ReContraster: Making Your Posters Stand Out with Regional Contrast

DGX agent

arXiv:2604.10442v1 Announce Type: new Abstract: Effective poster design requires rapidly capturing attention and clearly conveying messages. Inspired by the ``contrast effects'' principle, we propose

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Reducing Hallucination in Enterprise AI Workflows via Hybrid Utility Minimum Bayes Risk (HUMBR)

DGX agent

arXiv:2604.11141v1 Announce Type: new Abstract: Although LLMs drive automation, it is critical to ensure immense consideration for high-stakes enterprise workflows such as those involving legal matter

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

ReFEree: Reference-Free and Fine-Grained Method for Evaluating Factual Consistency in Real-World Code Summarization

DGX agent

arXiv:2604.10520v1 Announce Type: cross Abstract: As Large Language Models (LLMs) have become capable of generating long and descriptive code summaries, accurate and reliable evaluation of factual con

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Relational Preference Encoding in Looped Transformer Internal States

DGX agent

arXiv:2604.09870v1 Announce Type: cross Abstract: We investigate how looped transformers encode human preference in their internal iteration states. Using Ouro-2.6B-Thinking, a 2.6B-parameter looped t

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Relax: An Asynchronous Reinforcement Learning Engine for Omni-Modal Post-Training at Scale

DGX agent

arXiv:2604.11554v1 Announce Type: new Abstract: Reinforcement learning (RL) post-training has proven effective at unlocking reasoning, self-reflection, and tool-use capabilities in large language mode

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

ReplicateAnyScene: Zero-Shot Video-to-3D Composition via Textual-Visual-Spatial Alignment

DGX agent

arXiv:2604.10789v1 Announce Type: new Abstract: Humans exhibit an innate capacity to rapidly perceive and segment objects from video observations, and even mentally assemble them into structured 3D sc

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Respondology launches Respond to turn comments into competitive advantage

DGX agent

Respondology, the provider of a social media comment moderation and intelligence platform, today announced the launch of Respond, an artificial intelligence-powered platform that engages with social m

model-releasessiliconangle
14 Apr 2026
Model Releases

Rethinking Video Human-Object Interaction: Set Prediction over Time for Unified Detection and Anticipation

DGX agent

arXiv:2604.10397v1 Announce Type: cross Abstract: Video-based human-object interaction (HOI) understanding requires both detecting ongoing interactions and anticipating their future evolution. However

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Retrieval-Augmented Large Language Models for Evidence-Informed Guidance on Cannabidiol Use in Older Adults

DGX agent

arXiv:2604.09548v1 Announce Type: cross Abstract: Older adults commonly experience chronic conditions such as pain and sleep disturbances and may consider cannabidiol for symptom management. Safe use

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Revisiting the Scale Loss Function and Gaussian-Shape Convolution for Infrared Small Target Detection

DGX agent

arXiv:2604.09991v1 Announce Type: new Abstract: Infrared small target detection still faces two persistent challenges: training instability from non-monotonic scale loss functions, and inadequate spat

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

ReXSonoVQA: A Video QA Benchmark for Procedure-Centric Ultrasound Understanding

DGX agent

arXiv:2604.10916v1 Announce Type: cross Abstract: Ultrasound acquisition requires skilled probe manipulation and real-time adjustments. Vision-language models (VLMs) could enable autonomous ultrasound

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Rhizome OS-1: Rhizome's Semi-Autonomous Operating System for Small Molecule Drug Discovery

DGX agent

arXiv:2604.07512v2 Announce Type: replace Abstract: We present Rhizome OS-1, a semi-autonomous operating system for small molecule drug discovery in which multi-modal AI agents operate as a full multi

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

RISK: A Framework for GUI Agents in E-commerce Risk Management

DGX agent

arXiv:2509.21982v2 Announce Type: replace Abstract: E-commerce risk management requires aggregating diverse, deeply embedded web data through multi-step, stateful interactions, which traditional scrap

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

RiTeK: A Dataset for Large Language Models Complex Reasoning over Textual Knowledge Graphs in Medicine

DGX agent

arXiv:2410.13987v3 Announce Type: replace Abstract: Answering complex real-world questions in the medical domain often requires accurate retrieval from medical Textual Knowledge Graphs (medical TKGs),

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

RL-Driven Sustainable Land-Use Allocation for the Lake Malawi Basin

DGX agent

arXiv:2604.03768v2 Announce Type: replace Abstract: Unsustainable land-use practices in ecologically sensitive regions threaten biodiversity, water resources, and the livelihoods of millions. This pap

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

RL makes MLLMs see better than SFT

DGX agent

arXiv:2510.16333v2 Announce Type: replace Abstract: A dominant assumption in Multimodal Language Model (MLLM) research is that its performance is largely inherited from the LLM backbone, given its imm

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

RoboLab: A High-Fidelity Simulation Benchmark for Analysis of Task Generalist Policies

DGX agent

arXiv:2604.09860v1 Announce Type: cross Abstract: The pursuit of general-purpose robotics has yielded impressive foundation models, yet simulation-based benchmarking remains a bottleneck due to rapid

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Robust Adversarial Policy Optimization Under Dynamics Uncertainty

DGX agent

arXiv:2604.10974v1 Announce Type: new Abstract: Reinforcement learning (RL) policies often fail under dynamics that differ from training, a gap not fully addressed by domain randomization or existing

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

Robust Fair Disease Diagnosis in CT Images

DGX agent

arXiv:2604.09710v1 Announce Type: new Abstract: Automated diagnosis from chest CT has improved considerably with deep learning, but models trained on skewed datasets tend to perform unevenly across pa

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

RobustMedSAM: Degradation-Resilient Medical Image Segmentation via Robust Foundation Model Adaptation

DGX agent

arXiv:2604.09814v1 Announce Type: new Abstract: Medical image segmentation models built on Segment Anything Model (SAM) achieve strong performance on clean benchmarks, yet their reliability often degr

model-releasesarxiv-cs-cv
14 Apr 2026
← Previous
1…442443444445446…464
Next →