AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,280 results
14 Apr 2026

Observe Less, Understand More: Cost-aware Cross-scale Observation for Remote Sensing Understanding

Model ReleasesDGX agent

arXiv:2604.11415v1 Announce Type: new Abstract: Remote sensing understanding inherently requires multi-resolution observation, since different targets and application tasks demand different levels of

OccuBench: Evaluating AI Agents on Real-World Professional Tasks via Language World Models

Model ReleasesDGX agent

arXiv:2604.10866v1 Announce Type: new Abstract: AI agents are expected to perform professional work across hundreds of occupational domains (from emergency department triage to nuclear reactor safety

ODUTQA-MDC: A Task for Open-Domain Underspecified Tabular QA with Multi-turn Dialogue-based Clarification


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

arXiv:2604.10159v1 Announce Type: new Abstract: The advancement of large language models (LLMs) has enhanced tabular question answering (Tabular QA), yet they struggle with open-domain queries exhibit

Ollama and Google Gemma team is hosting an Ollama Gemma Day in Palo Alto tomorrow night (Wednesday, April 15th at 6pm). Lots of amazing spea…

Model ReleasesDGX agent

Ollama and Google Gemma team is hosting an Ollama Gemma Day in Palo Alto tomorrow night (Wednesday, April 15th at 6pm). Lots of amazing speakers from @GoogleDeepMind and @sgl_project / @radixark! Look

Ollama Max vs. Claude Code vs. ChatGPT Plan

Model ReleasesDGX agent

This Reddit thread from r/ollama compares three AI subscription/access options — Ollama Max (a paid Ollama tier), Claude Code (Anthropic's coding-focused offering), and a ChatGPT paid plan — likely ev

Omnimodal Dataset Distillation via High-order Proxy Alignment

Model ReleasesDGX agent

arXiv:2604.10666v1 Announce Type: cross Abstract: Dataset distillation compresses large-scale datasets into compact synthetic sets while preserving training performance, but existing methods are large

OmniScript: Towards Audio-Visual Script Generation for Long-Form Cinematic Video

Model ReleasesDGX agent

arXiv:2604.11102v1 Announce Type: new Abstract: Current multimodal large language models (MLLMs) have demonstrated remarkable capabilities in short-form video understanding, yet translating long-form

OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation

Model ReleasesDGX agent

arXiv:2604.11804v1 Announce Type: new Abstract: In this work, we study Human-Object Interaction Video Generation (HOIVG), which aims to synthesize high-quality human-object interaction videos conditio

On Harnessing Idle Compute at the Edge for Foundation Model Training

Model ReleasesDGX agent

arXiv:2512.22142v2 Announce Type: replace-cross Abstract: The foundation-model ecosystem remains highly centralized because training requires immense compute resources and is therefore largely limited

One of my passions is that education should be dispersed freely and as widely as possible, especially for technologies as dynamic and crucia…

Model ReleasesDGX agent

One of my passions is that education should be dispersed freely and as widely as possible, especially for technologies as dynamic and crucial as LLMs/AI. I'm proud to have friends who would disown me

Online Covariance Estimation in Averaged SGD: Improved Batch-Mean Rates and Minimax Optimality via Trajectory Regression

Model ReleasesDGX agent

arXiv:2604.10814v1 Announce Type: new Abstract: We study online covariance matrix estimation for Polyak--Ruppert averaged stochastic gradient descent (SGD). The online batch-means estimator of Zhu, Ch

Online Covariance Matrix Estimation in Sketched Newton Methods

Model ReleasesDGX agent

arXiv:2502.07114v2 Announce Type: replace-cross Abstract: Given the ubiquity of streaming data, online algorithms have been widely used for parameter estimation, with second-order methods particularly

Online Reasoning Video Object Segmentation

Model ReleasesDGX agent

arXiv:2604.11411v1 Announce Type: new Abstract: Reasoning video object segmentation predicts pixel-level masks in videos from natural-language queries that may involve implicit and temporally grounded

OOWM: Structuring Embodied Reasoning and Planning via Object-Oriented Programmatic World Modeling

Model ReleasesDGX agent

arXiv:2604.09580v1 Announce Type: new Abstract: Standard Chain-of-Thought (CoT) prompting empowers Large Language Models (LLMs) with reasoning capabilities, yet its reliance on linear natural language

Open Harness 🤝 Deployed Agents if you wanna use Claude, GLM5, and Codex in your deployed harness then you should be able to! deepagents dep…

Model ReleasesDGX agent

Open Harness 🤝 Deployed Agents if you wanna use Claude, GLM5, and Codex in your deployed harness then you should be able to! deepagents deploy has easy configs to let users customize their harness and

OpenAI launches GPT-5.4-Cyber model for vetted security pros

Model ReleasesDGX agent

OpenAI Group PBC today announced the launch of GPT-5.4-Cyber, a fine-tuned variant of its GPT-5.4 model designed for defensive cybersecurity work and also announced a significant expansion of its Trus

OpenAI rolls out GPT-5.4-Cyber, a fine-tuned GPT-5.4 variant for defensive cybersecurity use cases, to some participants of its Trusted Access for Cyber program (Rachel Metz/Bloomberg)

Model ReleasesDGX agent

Rachel Metz / Bloomberg: OpenAI rolls out GPT-5.4-Cyber, a fine-tuned GPT-5.4 variant for defensive cybersecurity use cases, to some participants of its Trusted Access for Cyber program — OpenAI is le

OpenAI's economic agenda is oddly socialist and wildly hypocritical, and is largely undermined by OpenAI's support for Republicans who attack welfare programs (Eric Levitz/Vox)

Model ReleasesDGX agent

Eric Levitz / Vox: OpenAI's economic agenda is oddly socialist and wildly hypocritical, and is largely undermined by OpenAI's support for Republicans who attack welfare programs — The AI company relea

Orthogonal Quadratic Complements for Vision Transformer Feed-Forward Networks

Model ReleasesDGX agent

arXiv:2604.09709v1 Announce Type: cross Abstract: Recent bilinear feed-forward replacements for vision transformers can substantially improve accuracy, but they often conflate two effects: stronger se

PAC-BENCH: Evaluating Multi-Agent Collaboration under Privacy Constraints

Model ReleasesDGX agent

arXiv:2604.11523v1 Announce Type: new Abstract: We are entering an era in which individuals and organizations increasingly deploy dedicated AI agents that interact and collaborate with other agents. H

Pando: Do Interpretability Methods Work When Models Won't Explain Themselves?

Model ReleasesDGX agent

arXiv:2604.11061v1 Announce Type: cross Abstract: Mechanistic interpretability is often motivated for alignment auditing, where a model's verbal explanations can be absent, incomplete, or misleading.

Panoptic Pairwise Distortion Graph

Model ReleasesDGX agent

arXiv:2604.11004v1 Announce Type: cross Abstract: In this work, we introduce a new perspective on comparative image assessment by representing an image pair as a structured composition of its regions.

PaperScope: A Multi-Modal Multi-Document Benchmark for Agentic Deep Research Across Massive Scientific Papers

Model ReleasesDGX agent

arXiv:2604.11307v1 Announce Type: new Abstract: Leveraging Multi-modal Large Language Models (MLLMs) to accelerate frontier scientific research is promising, yet how to rigorously evaluate such system

Para-B&B: Load-Balanced Deterministic Parallelization of Solving MIP

Model ReleasesDGX agent

arXiv:2604.09556v1 Announce Type: cross Abstract: Mixed-integer programming (MIP) extends linear programming by incorporating both continuous and integer decision variables, making it widely used in p

Parameter Efficient Fine-tuning for Domain-specific Gastrointestinal Disease Recognition

Model ReleasesDGX agent

arXiv:2604.10451v1 Announce Type: new Abstract: Despite recent advancements in the field of medical image analysis with the use of pretrained foundation models, the issue of distribution shifts betwee

ParseBench is the most comprehensive OCR benchmark for real-world enterprise documents: financial filings, contracts, insurance documents, a…

Model ReleasesDGX agent

ParseBench is the most comprehensive OCR benchmark for real-world enterprise documents: financial filings, contracts, insurance documents, and more. We evaluate across 5 dimensions that are present am

PepBenchmark: A Standardized Benchmark for Peptide Machine Learning

Model ReleasesDGX agent

arXiv:2604.10531v1 Announce Type: cross Abstract: Peptide therapeutics are widely regarded as the 'third generation' of drugs, yet progress in peptide Machine Learning (ML) are hindered by the absence

Persona Non Grata: Single-Method Safety Evaluation Is Incomplete for Persona-Imbued LLMs

Model ReleasesDGX agent

arXiv:2604.11120v1 Announce Type: new Abstract: Personality imbuing customizes LLM behavior, but safety evaluations almost always study prompt-based personas alone. We show this is incomplete: prompti

PhyMix: Towards Physically Consistent Single-Image 3D Indoor Scene Generation with Implicit--Explicit Optimization

Model ReleasesDGX agent

arXiv:2604.10125v1 Announce Type: new Abstract: Existing single-image 3D indoor scene generators often produce results that look visually plausible but fail to obey real-world physics, limiting their

Physics-informed AI Accelerated Retention Analysis of Ferroelectric Vertical NAND: From Day-Scale TCAD to Second-Scale Surrogate Model

Model ReleasesDGX agent

arXiv:2603.06881v2 Announce Type: replace-cross Abstract: Ferroelectric field-effect transistors (FeFET)-based vertical NAND (Fe-VNAND) has emerged as a promising candidate to overcome z-scaling limit

Pioneer Agent: Continual Improvement of Small Language Models in Production

Model ReleasesDGX agent

arXiv:2604.09791v1 Announce Type: new Abstract: Small language models are attractive for production deployment due to their low cost, fast inference, and ease of specialization. However, adapting them

Playing Along: Learning a Double-Agent Defender for Belief Steering via Theory of Mind

Model ReleasesDGX agent

arXiv:2604.11666v1 Announce Type: cross Abstract: As large language models (LLMs) become the engine behind conversational systems, their ability to reason about the intentions and states of their dial

Please Make it Sound like Human: Encoder-Decoder vs. Decoder-Only Transformers for AI-to-Human Text Style Transfer

Model ReleasesDGX agent

arXiv:2604.11687v1 Announce Type: new Abstract: AI-generated text has become common in academic and professional writing, prompting research into detection methods. Less studied is the reverse: system

Point2Pose: Occlusion-Recovering 6D Pose Tracking and 3D Reconstruction for Multiple Unknown Objects Via 2D Point Trackers

Model ReleasesDGX agent

arXiv:2604.10415v1 Announce Type: new Abstract: We present Point2Pose, a model-free method for causal 6D pose tracking of multiple rigid objects from monocular RGB-D video. Initialized only from spars

PokeRL: Reinforcement Learning for Pokemon Red

Model ReleasesDGX agent

arXiv:2604.10812v1 Announce Type: new Abstract: Pokemon Red is a long-horizon JRPG with sparse rewards, partial observability, and quirky control mechanics that make it a challenging benchmark for rei

Policy-Guided Threat Hunting: An LLM enabled Framework with Splunk SOC Triage

Model ReleasesDGX agent

arXiv:2603.23966v3 Announce Type: replace-cross Abstract: With frequently evolving Advanced Persistent Threats (APTs) in cyberspace, traditional security solutions approaches have become inadequate fo

Polyglot Teachers: Evaluating Language Models for Multilingual Synthetic Data Generation

Model ReleasesDGX agent

arXiv:2604.11290v1 Announce Type: new Abstract: Synthesizing supervised finetuning (SFT) data from language models (LMs) to teach smaller models multilingual tasks has become increasingly common. Howe

Powerful Training-Free Membership Inference Against Autoregressive Language Models

Model ReleasesDGX agent

arXiv:2601.12104v2 Announce Type: replace-cross Abstract: Fine-tuned language models pose significant privacy risks, as they may memorize and expose sensitive information from their training data. Mem

Precision Synthesis of Multi-Tracer PET via VLM-Modulated Rectified Flow for Stratifying Mild Cognitive Impairment

Model ReleasesDGX agent

arXiv:2604.11176v1 Announce Type: new Abstract: The biological definition of Alzheimer's disease (AD) relies on multi-modal neuroimaging, yet the clinical utility of positron emission tomography (PET)

ProGAL-VLA: Grounded Alignment through Prospective Reasoning in Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2604.09824v1 Announce Type: cross Abstract: Vision language action (VLA) models enable generalist robotic agents but often exhibit language ignorance, relying on visual shortcuts and remaining i

Progressive Multimodal Interaction Network for Reliable Quantification of Fish Feeding Intensity in Aquaculture

Model ReleasesDGX agent

arXiv:2506.14170v3 Announce Type: replace-cross Abstract: Accurate quantification of fish feeding intensity is crucial for precision feeding in aquaculture, as it directly affects feed utilization and

PSF-Med: Measuring and Explaining Paraphrase Sensitivity in Medical Vision Language Models

Model ReleasesDGX agent

arXiv:2602.21428v2 Announce Type: replace Abstract: Medical Vision Language Models (VLMs) can change their answers when clinicians rephrase the same question, a failure mode that threatens deployment

Quantitative Introspection in Language Models: Tracking Emotive States Across Conversation

Model ReleasesDGX agent

arXiv:2603.18893v2 Announce Type: replace Abstract: Tracking the internal states of large language models across conversations is important for safety, interpretability, and model welfare, yet current

Quantization Dominates Rank Reduction for KV-Cache Compression

Model ReleasesDGX agent

arXiv:2604.11501v1 Announce Type: cross Abstract: We compare two strategies for compressing the KV cache in transformer inference: rank reduction (discard dimensions) and quantization (keep all dimens

Radiology Report Generation for Low-Quality X-Ray Images

Model ReleasesDGX agent

arXiv:2604.10188v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have significantly advanced automated Radiology Report Generation (RRG). However, existing methods implicitly assume high-

RationalRewards: Reasoning Rewards Scale Visual Generation Both Training and Test Time

Model ReleasesDGX agent

arXiv:2604.11626v1 Announce Type: new Abstract: Most reward models for visual generation reduce rich human judgments to a single unexplained score, discarding the reasoning that underlies preference.

RCBSF: A Multi-Agent Framework for Automated Contract Revision via Stackelberg Game

Model ReleasesDGX agent

arXiv:2604.10740v1 Announce Type: new Abstract: Despite the widespread adoption of Large Language Models (LLMs) in Legal AI, their utility for automated contract revision remains impeded by hallucinat

Reasoning as Gradient: Scaling MLE Agents Beyond Tree Search

Model ReleasesDGX agent

arXiv:2603.01692v3 Announce Type: replace-cross Abstract: LLM-based agents for machine learning engineering (MLE) predominantly rely on tree search, a form of gradient-free optimization that uses scal

ReContraster: Making Your Posters Stand Out with Regional Contrast

Model ReleasesDGX agent

arXiv:2604.10442v1 Announce Type: new Abstract: Effective poster design requires rapidly capturing attention and clearly conveying messages. Inspired by the ``contrast effects'' principle, we propose

Reducing Hallucination in Enterprise AI Workflows via Hybrid Utility Minimum Bayes Risk (HUMBR)

Model ReleasesDGX agent

arXiv:2604.11141v1 Announce Type: new Abstract: Although LLMs drive automation, it is critical to ensure immense consideration for high-stakes enterprise workflows such as those involving legal matter

ReFEree: Reference-Free and Fine-Grained Method for Evaluating Factual Consistency in Real-World Code Summarization

Model ReleasesDGX agent

arXiv:2604.10520v1 Announce Type: cross Abstract: As Large Language Models (LLMs) have become capable of generating long and descriptive code summaries, accurate and reliable evaluation of factual con

Relational Preference Encoding in Looped Transformer Internal States

Model ReleasesDGX agent

arXiv:2604.09870v1 Announce Type: cross Abstract: We investigate how looped transformers encode human preference in their internal iteration states. Using Ouro-2.6B-Thinking, a 2.6B-parameter looped t

Relax: An Asynchronous Reinforcement Learning Engine for Omni-Modal Post-Training at Scale

Model ReleasesDGX agent

arXiv:2604.11554v1 Announce Type: new Abstract: Reinforcement learning (RL) post-training has proven effective at unlocking reasoning, self-reflection, and tool-use capabilities in large language mode

ReplicateAnyScene: Zero-Shot Video-to-3D Composition via Textual-Visual-Spatial Alignment

Model ReleasesDGX agent

arXiv:2604.10789v1 Announce Type: new Abstract: Humans exhibit an innate capacity to rapidly perceive and segment objects from video observations, and even mentally assemble them into structured 3D sc

Respondology launches Respond to turn comments into competitive advantage

Model ReleasesDGX agent

Respondology, the provider of a social media comment moderation and intelligence platform, today announced the launch of Respond, an artificial intelligence-powered platform that engages with social m

Rethinking Video Human-Object Interaction: Set Prediction over Time for Unified Detection and Anticipation

Model ReleasesDGX agent

arXiv:2604.10397v1 Announce Type: cross Abstract: Video-based human-object interaction (HOI) understanding requires both detecting ongoing interactions and anticipating their future evolution. However

Retrieval-Augmented Large Language Models for Evidence-Informed Guidance on Cannabidiol Use in Older Adults

Model ReleasesDGX agent

arXiv:2604.09548v1 Announce Type: cross Abstract: Older adults commonly experience chronic conditions such as pain and sleep disturbances and may consider cannabidiol for symptom management. Safe use

Revisiting the Scale Loss Function and Gaussian-Shape Convolution for Infrared Small Target Detection

Model ReleasesDGX agent

arXiv:2604.09991v1 Announce Type: new Abstract: Infrared small target detection still faces two persistent challenges: training instability from non-monotonic scale loss functions, and inadequate spat

ReXSonoVQA: A Video QA Benchmark for Procedure-Centric Ultrasound Understanding

Model ReleasesDGX agent

arXiv:2604.10916v1 Announce Type: cross Abstract: Ultrasound acquisition requires skilled probe manipulation and real-time adjustments. Vision-language models (VLMs) could enable autonomous ultrasound

Rhizome OS-1: Rhizome's Semi-Autonomous Operating System for Small Molecule Drug Discovery

Model ReleasesDGX agent

arXiv:2604.07512v2 Announce Type: replace Abstract: We present Rhizome OS-1, a semi-autonomous operating system for small molecule drug discovery in which multi-modal AI agents operate as a full multi

← Previous
1…354355356357358…372
Next →