AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
All
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,315 results
Model Releases

Open-TQ-Metal: Fused Compressed-Domain Attention for Long-Context LLM Inference on Apple Silicon

DGX agent

arXiv:2604.16957v1 Announce Type: new Abstract: We present Open-TQ-Metal, the first implementation of fused compressed-domain attention on Apple Silicon, enabling 128K-context inference for Llama 3.1

model-releasesarxiv-cs-lg
21 Apr 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

OpenAI launches ChatGPT Images 2.0, Codex Labs developer training service

DGX agent

OpenAI Group PBC today launched ChatGPT Images 2.0, an upgraded version of the image generator built into its popular chatbot. The company also debuted a new technical training service called Codex La

model-releasessiliconangle
21 Apr 2026
Model Releases

OPeRA: A Dataset of Observation, Persona, Rationale, and Action for Evaluating LLMs on Human Online Shopping Behavior Simulation

DGX agent

arXiv:2506.05606v5 Announce Type: replace Abstract: Can large language models (LLMs) accurately simulate the next web action of a specific user? While LLMs have shown promising capabilities in generat

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

OptiMVMap: Offline Vectorized Map Construction via Optimal Multi-vehicle Perspectives

DGX agent

arXiv:2604.17135v1 Announce Type: new Abstract: Offline vectorized maps constitute critical infrastructure for high-precision autonomous driving and mapping services. Existing approaches rely predomin

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

OptunaHub: A Platform for Black-Box Optimization

DGX agent

arXiv:2510.02798v2 Announce Type: replace Abstract: Black-box optimization (BBO) underpins advances in domains such as AutoML and Materials Informatics, yet implementations of algorithms and benchmark

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Overcoming Selection Bias in Statistical Studies With Amortized Bayesian Inference

DGX agent

arXiv:2604.18319v1 Announce Type: cross Abstract: Selection bias arises when the probability that an observation enters a dataset depends on variables related to the quantities of interest, leading to

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

PAC-Bayes Bounds for Gibbs Posteriors via Singular Learning Theory

DGX agent

arXiv:2604.17219v1 Announce Type: cross Abstract: We derive explicit non-asymptotic PAC-Bayes generalization bounds for Gibbs posteriors, that is, data-dependent distributions over model parameters ob

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Parallel Test-Time Scaling for Latent Reasoning Models

DGX agent

arXiv:2510.07745v4 Announce Type: replace Abstract: Parallel test-time scaling (TTS) is a pivotal approach for enhancing large language models (LLMs), typically by sampling multiple token-based chains

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

ParseBench is the first benchmark to include VLM chart understanding 📊📈📉 over enterprise documents. 🟠 Existing benchmarks (ChartQA, Char…

DGX agent

ParseBench is the first benchmark to include VLM chart understanding 📊📈📉 over enterprise documents. 🟠 Existing benchmarks (ChartQA, ChartXiv) test over charts specifically and not the chart's inclusio

model-releasesjerry-liu--x
21 Apr 2026
Model Releases

PBSBench: A Multi-Level Vision-Language Framework and Benchmark for Hematopathology Whole Slide Image Interpretation

DGX agent

arXiv:2604.17570v1 Announce Type: new Abstract: Peripheral Blood Smear (PBS) is a critical microscopic examination in hematopathology that yields whole-slide imaging (WSI). Unlike solid tissue patholo

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

PDDL-Mind: Large Language Models are Capable on Belief Reasoning with Reliable State Tracking

DGX agent

arXiv:2604.17819v1 Announce Type: new Abstract: Large language models (LLMs) perform substantially below human level on existing theory-of-mind (ToM) benchmarks, even when augmented with chain-of-thou

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Penny Wise, Pixel Foolish: Bypassing Price Constraints in Multimodal Agents via Visual Adversarial Perturbations

DGX agent

arXiv:2604.16515v1 Announce Type: new Abstract: The rapid proliferation of Multimodal Large Language Models (MLLMs) has enabled mobile agents to execute high-stakes financial transactions, but their a

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

people compared GPT-5.4's solution of erdos #1196 to alphago's move 37 but i think a tighter analogy is to alphago's unusual (at the time) p…

DGX agent

people compared GPT-5.4's solution of erdos #1196 to alphago's move 37 but i think a tighter analogy is to alphago's unusual (at the time) preferences for the 3-3 opening and early 3-3 invasion agains

model-releasesemad-mostaque--x
21 Apr 2026
Model Releases

PersonalHomeBench: Evaluating Agents in Personalized Smart Homes

DGX agent

arXiv:2604.16813v1 Announce Type: cross Abstract: Agentic AI systems are rapidly advancing toward real-world applications, yet their readiness in complex and personalized environments remains insuffic

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

PFDelta: A Benchmark Dataset for Power Flow under Load, Generation, and Topology Variations

DGX agent

arXiv:2510.22048v3 Announce Type: replace Abstract: Power flow (PF) calculations are the backbone of real-time grid operations, across workflows such as contingency analysis (where repeated PF evaluat

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Physics-Informed Causal MDPs for Sequential Constraint Repair in Engineering Simulation Pipelines

DGX agent

arXiv:2604.17910v1 Announce Type: cross Abstract: Off-policy learning in constrained MDPs with large binary state spaces faces a fundamental tension: causal identification of transition dynamics requi

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Physics-Informed Neural Networks: A Didactic Derivation of the Complete Training Cycle

DGX agent

arXiv:2604.18481v1 Announce Type: cross Abstract: This paper is a step-by-step, self-contained guide to the complete training cycle of a Physics-Informed Neural Network (PINN) -- a topic that existing

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

PiCa: Parameter-Efficient Fine-Tuning with Column Space Projection

DGX agent

arXiv:2505.20211v3 Announce Type: replace Abstract: Fine-tuning large foundation models is essential for building expert models tailored to specialized tasks and domains, but fully updating billions o

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

PlanViz: Evaluating Planning-Oriented Image Generation and Editing for Computer-Use Tasks

DGX agent

arXiv:2602.06663v2 Announce Type: replace Abstract: Unified multimodal models (UMMs) have shown impressive capabilities in generating natural images and supporting multimodal reasoning. However, their

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Please don’t trust your chatbot for medical advice. 🙏 Remember how I used to say that large language models are “frequently wrong, never in…

DGX agent

Please don’t trust your chatbot for medical advice. 🙏 Remember how I used to say that large language models are “frequently wrong, never in doubt”, and how I warned three years ago on 60 Minutes that

model-releasesgary-marcus--x
21 Apr 2026
Model Releases

Please refuse to answer me! Mitigating Over-Refusal in Large Language Models via Adaptive Contrastive Decoding

DGX agent

arXiv:2604.17132v1 Announce Type: new Abstract: Safety-aligned large language models (LLMs) often generate refusal responses to harmless queries due to the over-refusal problem. However, existing meth

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Polysemantic Experts, Monosemantic Paths: Routing as Control in MoEs

DGX agent

arXiv:2604.17837v1 Announce Type: cross Abstract: An LLM's residual stream is both state and instruction: it encodes the current context and determines the next transformation. We introduce a paramete

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

PPEDCRF: Dynamic-CRF-Guided Selective Perturbation for Background-Based Location Privacy in Video Sequences

DGX agent

arXiv:2604.17163v1 Announce Type: new Abstract: We propose PPEDCRF, a calibrated selective perturbation framework that protects background-based location privacy in released video frames against galle

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Precise Debugging Benchmark: Is Your Model Debugging or Regenerating?

DGX agent

arXiv:2604.17338v1 Announce Type: cross Abstract: Unlike code completion, debugging requires localizing faults and applying targeted edits. We observe that frontier LLMs often regenerate correct but o

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Predicting LLM Compression Degradation from Spectral Statistics

DGX agent

arXiv:2604.18085v1 Announce Type: new Abstract: Matrix-level low-rank compression is a promising way to reduce the cost of large language models, but running compression and evaluating the resulting m

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Pref-GRPO: Pairwise Preference Reward-based GRPO for Stable Text-to-Image Reinforcement Learning

DGX agent

arXiv:2508.20751v2 Announce Type: replace Abstract: Recent advancements highlight the importance of GRPO-based reinforcement learning methods and benchmarking in enhancing text-to-image (T2I) generati

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

PrefixMemory-Tuning: Modernizing Prefix-Tuning by Decoupling the Prefix from Attention

DGX agent

arXiv:2506.13674v3 Announce Type: replace Abstract: Parameter-Efficient Fine-Tuning (PEFT) methods have become crucial for rapidly adapting large language models (LLMs) to downstream tasks. Prefix-Tun

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Prior-Fitted Functional Flow: In-Context Generative Models for Pharmacokinetics

DGX agent

arXiv:2604.17670v1 Announce Type: new Abstract: We introduce Prior-Fitted Functional Flows, a generative foundation model for pharmacokinetics that enables zero-shot population synthesis and individua

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

PRISM: Probing Reasoning, Instruction, and Source Memory in LLM Hallucinations

DGX agent

arXiv:2604.16909v1 Announce Type: new Abstract: As large language models (LLMs) evolve from conversational assistants into agents capable of handling complex tasks, they are increasingly deployed in h

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

ProfVLM: A lightweight video-language model for multi-view proficiency estimation

DGX agent

arXiv:2509.26278v4 Announce Type: replace-cross Abstract: Most existing approaches formulate action quality assessment and skill proficiency estimation as discriminative prediction tasks, typically pr

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Program Structure-aware Language Models: Targeted Software Testing beyond Textual Semantics

DGX agent

arXiv:2604.17715v1 Announce Type: cross Abstract: Recent advances in large language models for test case generation have improved branch coverage via prompt-engineered mutations. However, they still l

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Progressive Online Video Understanding with Evidence-Aligned Timing and Transparent Decisions

DGX agent

arXiv:2604.18459v1 Announce Type: new Abstract: Visual agents operating in the wild must respond to queries precisely when sufficient evidence first appears in a video stream, a critical capability th

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Prompting Foundation Models for Zero-Shot Ship Instance Segmentation in SAR Imagery

DGX agent

arXiv:2604.17920v1 Announce Type: new Abstract: Synthetic Aperture Radar (SAR) plays a critical role in maritime surveillance, yet deep learning for SAR analysis is limited by the lack of pixel-level

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

ProTrain: Efficient LLM Training via Memory-Aware Techniques

DGX agent

arXiv:2406.08334v2 Announce Type: replace-cross Abstract: Memory pressure has emerged as a dominant constraint in scaling the training of large language models (LLMs), particularly in resource-constra

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Prune, Interpret, Evaluate: A Cross-Layer Transcoder-Native Framework for Efficient Circuit Discovery via Feature Attribution

DGX agent

arXiv:2604.16889v1 Announce Type: new Abstract: Existing feature-interpretation pipelines typically operate on uniformly sampled units, but only a small fraction of cross-layer transcoder (CLT) featur

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Pseudo2Real: Task Arithmetic for Pseudo-Label Correction in Automatic Speech Recognition

DGX agent

arXiv:2510.08047v2 Announce Type: replace-cross Abstract: Robust ASR under domain shift is crucial because real-world systems encounter unseen accents and domains with limited labeled data. Although p

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

PTLD: Sim-to-real Privileged Tactile Latent Distillation for Dexterous Manipulation

DGX agent

arXiv:2603.04531v2 Announce Type: replace Abstract: Tactile dexterous manipulation is essential to automating complex household tasks, yet learning effective control policies remains a challenge. Whil

model-releasesarxiv-cs-ro
21 Apr 2026
Model Releases

Pulse Shape Discrimination Algorithms: Survey and Benchmark

DGX agent

arXiv:2508.02750v2 Announce Type: replace Abstract: This review presents a comprehensive survey and benchmark of pulse shape discrimination (PSD) algorithms for radiation detection, classifying nearly

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

QU-NLP at QIAS 2026: Multi-Stage QLoRA Fine-Tuning for Arabic Islamic Inheritance Reasoning

DGX agent

arXiv:2604.16396v1 Announce Type: new Abstract: Islamic inheritance law (ilm al-mawar{i}th) presents a challenging domain for evaluating large language models' structured reasoning capabilities, requi

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

ReasonEmbed: Enhanced Text Embeddings for Reasoning-Intensive Document Retrieval

DGX agent

arXiv:2510.08252v2 Announce Type: replace-cross Abstract: In this paper, we introduce ReasonEmbed, a novel text embedding model developed for reasoning-intensive document retrieval. Our work includes

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

ReCap: Lightweight Referential Grounding for Coherent Story Visualization

DGX agent

arXiv:2604.18575v1 Announce Type: new Abstract: Story Visualization aims to generate a sequence of images that faithfully depicts a textual narrative that preserve character identity, spatial configur

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

ReCoQA: A Benchmark for Tool-Augmented and Multi-Step Reasoning in Real Estate Question and Answering

DGX agent

arXiv:2604.17944v1 Announce Type: new Abstract: Developing agents capable of navigating fragmented, multi-source information remains challenging, primarily due to the scarcity of benchmarks reflecting

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

ReFineVLA: Multimodal Reasoning-Aware Generalist Robotic Policies via Teacher-Guided Fine-Tuning

DGX agent

arXiv:2604.17800v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have gained much attention from the research community thanks to their strength in translating multimodal observat

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

REFLEX: Self-Refining Explainable Fact-Checking via Verdict-Anchored Style Control

DGX agent

arXiv:2511.20233v3 Announce Type: replace Abstract: The prevalence of fake news on social media demands automated fact-checking systems to provide accurate verdicts with faithful explanations. However

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

ReflexiCoder: Teaching Large Language Models to Self-Reflect on Generated Code and Self-Correct It via Reinforcement Learning

DGX agent

arXiv:2603.05863v2 Announce Type: replace Abstract: While Large Language Models (LLMs) have revolutionized code generation, standard ``System 1'' approaches that generate solutions in a single forward

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Reinforced Efficient Reasoning via Semantically Diverse Exploration

DGX agent

arXiv:2601.05053v2 Announce Type: replace-cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has proven effective in enhancing the reasoning of large language models (LLMs). Monte C

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Reliability-Aware Adaptive Self-Consistency for Efficient Sampling in LLM Reasoning

DGX agent

arXiv:2601.02970v2 Announce Type: replace Abstract: Self-Consistency improves reasoning reliability through multi-sample aggregation, but incurs substantial inference cost. Adaptive self-consistency m

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Representation Before Training: A Fixed-Budget Benchmark for Generative Medical Event Models

DGX agent

arXiv:2604.16775v1 Announce Type: new Abstract: Every prediction from a generative medical event model is bounded by how clinical events are tokenized, yet input representation is rarely isolated from

model-releasesarxiv-cs-lg
21 Apr 2026
← Previous
1…411412413414415…465
Next →