AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,612 results
Model Releases

Njord: A Probabilistic Graph Neural Network for Ensemble Ocean Forecasting

DGX agent

arXiv:2605.15470v1 Announce Type: new Abstract: Ocean dynamics are inherently chaotic, yet existing machine learning ocean models produce only deterministic forecasts. We introduce Njord, a probabilis

model-releasesarxiv-cs-lg
18 May 2026
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Offline Semantic Guidance for Efficient Vision-Language-Action Policy Distillation

DGX agent

arXiv:2605.16241v1 Announce Type: cross Abstract: Billion-parameter Vision-Language-Action (VLA) policies have recently shown impressive performance in robotic manipulation, yet their size and inferen

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

OgBench: A Framework for Evaluating Graph Neural Networks on Omics Data

DGX agent

arXiv:2605.15511v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) have become the dominant framework for inductive graph-level learning. Yet most benchmarks focus on the regime n gg p, wher

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

okay this is going kinda viral and tbh my original text was kind of messy, so here's a second pass with the help of Claude: -- Implement <SP…

DGX agent

okay this is going kinda viral and tbh my original text was kind of messy, so here's a second pass with the help of Claude: -- Implement <SPEC>. As you work maintain a running implementation-notes.htm

model-releasesthariq--x
18 May 2026
Model Releases

On RGB-TIR Stereo Calibration under Extreme Resolution Asymmetry

DGX agent

arXiv:2605.15860v1 Announce Type: new Abstract: Accurate geometric calibration of RGB-thermal infrared (TIR) stereo camera systems is essential for multimodal building envelope analysis, yet remains c

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

One thing to watch for with Claude & GPT is that the models expose too much irrelevant history in their outputs. Slides are given footers sa…

DGX agent

One thing to watch for with Claude & GPT is that the models expose too much irrelevant history in their outputs. Slides are given footers saying things like 'Better, more targeted version' if you aske

model-releasesethan-mollick--x
18 May 2026
Model Releases

Optimizing LLM Inference: Fluid-Guided Online Scheduling with Memory Constraints

DGX agent

arXiv:2504.11320v3 Announce Type: replace-cross Abstract: Large language models now serve millions of users daily, with providers incurring costs exceeding $700,000 per day. Each request requires toke

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

PAGER: Bridging the Semantic-Execution Gap in Point-Precise Geometric GUI Control

DGX agent

arXiv:2605.15963v1 Announce Type: new Abstract: Large vision-language models have significantly advanced GUI agents, enabling executable interaction across web, mobile, and desktop interfaces. Yet the

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Painless Activation Steering: An Automated, Lightweight Approach for Post-Training Large Language Models

DGX agent

arXiv:2509.22739v3 Announce Type: replace-cross Abstract: Language models (LMs) are typically post-trained for desired capabilities and behaviors via weight-based or prompt-based steering, but the for

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

PBT-Bench: Benchmarking AI Agents on Property-Based Testing

DGX agent

arXiv:2605.15229v1 Announce Type: cross Abstract: Existing code benchmarks measure whether an agent can produce any test that reproduces a known bug, or whether it can produce a patch that fixes a des

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

PDRNN: Modular Data-driven Pedestrian Dead Reckoning on Loosely Coupled Radio- and Inertial-Signalstreams

DGX agent

arXiv:2605.15252v1 Announce Type: cross Abstract: Modern pedestrian dead reckoning (PDR) systems rely on fusing noisy and biased estimates of position, velocity, and calibrated orientation derived fro

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

PerfCodeBench: Benchmarking LLMs for System-Level High-Performance Code Optimization

DGX agent

arXiv:2605.15222v1 Announce Type: cross Abstract: Large language models (LLMs) can often generate functionally correct code, but their ability to produce efficient implementations for performance-crit

model-releasesarxiv-cs-cl
18 May 2026
Model Releases

Perforated Neural Networks for Keyword Spotting

DGX agent

arXiv:2605.15647v1 Announce Type: new Abstract: Edge machine learning presents a unique set of constraints not encountered in cloud-scale model deployment: strict memory budgets, limited compute, and

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

PhyDetEx: Detecting and Explaining the Physical Plausibility of T2V Models

DGX agent

arXiv:2512.01843v2 Announce Type: replace Abstract: Driven by the growing capacity and training scale, Text-to-Video (T2V) generation models have recently achieved substantial progress in video qualit

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

Position: Early-Stage Quality Assurance in Annotation Pipelines Is More Cost-Effective Than Late-Stage Validation

DGX agent

arXiv:2605.15714v1 Announce Type: cross Abstract: This position paper argues that the machine learning community should prioritize early-stage quality assurance in annotation pipelines over the prevai

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Position: Ideas Should be the Center of Machine Learning Research

DGX agent

arXiv:2605.15253v1 Announce Type: new Abstract: Machine learning research increasingly bifurcates into two disconnected modes: benchmark-driven engineering that prioritizes metrics over understanding,

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

Probabilistic Dating of Historical Manuscripts via Evidential Deep Regression on Visual Script Features

DGX agent

arXiv:2605.06475v1 Announce Type: cross Abstract: We introduce a probabilistic approach for dating historical manuscript pages from visual features alone. Instead of aggregating centuries into classes

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

Prompting Amazon Nova 2 for content moderation

DGX agent

In this post, you learn how to prompt Amazon Nova 2 Lite for content moderation using structured and free-form approaches, grounded in the MLCommons AILuminate Assessment Standard. The prompting techn

model-releasesaws-ml-blog
18 May 2026
Model Releases

Quantization Undoes Alignment: Bias Emergence in Compressed LLMs Across Models and Precision Levels

DGX agent

arXiv:2605.15208v1 Announce Type: cross Abstract: Large Language Models are routinely compressed via post-training quantization to reduce inference costs and memory footprint for cloud and edge deploy

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Quantum Feature Pyramid Gating for Seismic Image Segmentation

DGX agent

arXiv:2605.15370v1 Announce Type: cross Abstract: Accurate salt-body delineation is essential for seismic interpretation because salt structures distort wave propagation, complicate velocity-model bui

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

🚀🚀Qwen3.7 Preview lands on Arena ! Here come Qwen3.7-Max-Preview & Qwen3.7-Plus-Preview. Alibaba now #6 lab in Text, #5 in Vision.⚡️⚡️ Can…

DGX agent

🚀🚀Qwen3.7 Preview lands on Arena ! Here come Qwen3.7-Max-Preview & Qwen3.7-Plus-Preview. Alibaba now #6 lab in Text, #5 in Vision.⚡️⚡️ Can't wait to release Qwen3.7 series models!Stay tuned! @arena Qw

model-releasesqwen--x
18 May 2026
Model Releases

RapidUn: Influence-Driven Parameter Reweighting for Efficient Large Language Model Unlearning

DGX agent

arXiv:2512.04457v2 Announce Type: replace Abstract: Removing specific data influence from large language models (LLMs) remains challenging, as retraining is costly and existing approximate unlearning

model-releasesarxiv-cs-cl
18 May 2026
Model Releases

RAR: Retrieving And Ranking Augmented MLLMs for Visual Recognition

DGX agent

arXiv:2403.13805v2 Announce Type: replace-cross Abstract: CLIP (Contrastive Language-Image Pre-training) uses contrastive learning from noise image-text pairs to excel at recognizing a wide array of c

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Reducing the Safety Tax in LLM Safety Alignment with On-Policy Self-Distillation

DGX agent

arXiv:2605.15239v1 Announce Type: new Abstract: Safety alignment often improves robustness to harmful queries at the cost of reasoning ability, a tradeoff known as the safety tax. A common cause is di

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

Registers Matter for Pixel-Space Diffusion Transformers

DGX agent

arXiv:2605.16147v1 Announce Type: new Abstract: Vision Transformers (ViTs) are known to exhibit high-norm patch-token outliers that degrade feature map quality, a problem effectively mitigated by exti

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

Reinforcement learning for adaptive interior point methods in convex quadratic programming

DGX agent

arXiv:2509.07404v2 Announce Type: replace-cross Abstract: Quadratic programming is a workhorse of modern nonlinear optimization, control, and data science. Although regularized methods offer convergen

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

Representation Without Reward: A JEPA Audit for LLM Fine-Tuning

DGX agent

arXiv:2605.15394v1 Announce Type: cross Abstract: Joint-embedding predictive architectures (JEPAs) propose that a model should learn more useful abstractions when trained to predict latent representat

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Retrieval-Augmented Large Language Models for Schema-Constrained Clinical Information Extraction

DGX agent

arXiv:2605.15467v1 Announce Type: cross Abstract: Conversational nurse-patient transcripts contain actionable observations, but converting these transcripts into structured representations at scale re

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

RoadmapBench: Evaluating Long-Horizon Agentic Software Development Across Version Upgrades

DGX agent

arXiv:2605.15846v1 Announce Type: cross Abstract: Coding agents are increasingly deployed in real software development, where a single version iteration requires months of coordinated work across many

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

RTL-BenchMT: Dynamic Maintenance of RTL Generation Benchmark Through Agent-Assisted Analysis and Revision

DGX agent

arXiv:2605.15537v1 Announce Type: new Abstract: This paper introduces RTL-BenchMT, an agentic framework for dynamically maintaining RTL generation benchmarks. Large Language Models (LLMs) assisted aut

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Rule2DRC: Benchmarking LLM Agents for DRC Script Synthesis with Execution-Guided Test Generation

DGX agent

arXiv:2605.15669v1 Announce Type: new Abstract: Manufacturable chip layouts must satisfy thousands of geometry-based design rules, and design rule checking (DRC) enforces them by running executable DR

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

Run Claude Managed Agents with Vercel Sandbox

DGX agent

This article describes how to run Claude's managed agents within Vercel's Sandbox environment, enabling developers to execute AI agent workloads on Vercel's infrastructure. The integration allows user

model-releasesvercel-blog
18 May 2026
Model Releases

Runtime-Orchestrated Second-Order Optimization for Scalable LLM Training

DGX agent

arXiv:2605.16184v1 Announce Type: cross Abstract: Second-order methods offer an attractive path toward more sample-efficient LLM training, but their practical use is often blocked by the systems cost

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

SaaS-Bench: Can Computer-Use Agents Leverage Real-World SaaS to Solve Professional Workflows?

DGX agent

arXiv:2605.15777v1 Announce Type: new Abstract: Computer-Using Agents (CUAs) are rapidly extending large language models (LLMs) beyond text-based reasoning toward action execution in more complex envi

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

SARVLM: A Vision Language Foundation Model for Semantic Understanding in SAR Imagery

DGX agent

arXiv:2510.22665v3 Announce Type: replace-cross Abstract: Synthetic Aperture Radar (SAR) is a critical imaging modality due to its all-weather operational capability. Although recent advances in self-

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

SDOF: Taming the Alignment Tax in Multi-Agent Orchestration with State-Constrained Dispatch

DGX agent

arXiv:2605.15204v1 Announce Type: new Abstract: Multi-agent orchestration frameworks such as LangChain, LangGraph, and CrewAI route tasks through graph-based pipelines but do not enforce the stage con

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Searching on a Budget: HW-NAS with 10 Latency Probes

DGX agent

arXiv:2504.00663v2 Announce Type: replace Abstract: Existing hardware-aware NAS (HW-NAS) methods typically assume access to precise information circa the target device, either via analytical approxima

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

SemanticOpt: Towards LLM-Based Semantic Black-Box Optimization

DGX agent

arXiv:2510.25404v3 Announce Type: replace-cross Abstract: Optimizing an experimental system can be extremely challenging when each experiment is expensive, time-consuming, or difficult to perform. Exi

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

SGR: A Stepwise Reasoning Framework for LLMs with External Subgraph Generation

DGX agent

arXiv:2605.16117v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated strong capabilities across diverse NLP applications, such as translation, text generation, and question a

model-releasesarxiv-cs-cl
18 May 2026
Model Releases

ShopGym: An Integrated Framework for Realistic Simulation and Scalable Benchmarking of E-Commerce Web Agents

DGX agent

arXiv:2605.16116v1 Announce Type: new Abstract: Developing and evaluating e-commerce web agents requires environments that preserve meaningful task structure while enabling controllable, reproducible,

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

SkillSmith: Compiling Agent Skills into Boundary-Guided Runtime Interfaces

DGX agent

arXiv:2605.15215v1 Announce Type: new Abstract: Recently, skills have been widely adopted in large language model (LLM)-based agent systems across various domains. In existing frameworks, skills are t

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

SkyLink: A Large Vision-Language Model Driven Re-ranking Framework for Cross-View UAV geolocalization

DGX agent

arXiv:2603.08063v3 Announce Type: replace Abstract: Cross-view UAV geolocalization is fundamentally a challenging large-scale image retrieval task, aiming to determine the geographic coordinates of Un

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

Skyra: AI-Generated Video Detection via Grounded Artifact Reasoning

DGX agent

arXiv:2512.15693v2 Announce Type: replace Abstract: The misuse of AI-driven video generation technologies has raised serious social concerns, highlighting the urgent need for reliable AI-generated vid

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

SMMBench: A Benchmark for Source-Distributed Multimodal Agent Memory

DGX agent

arXiv:2605.15710v1 Announce Type: new Abstract: Existing benchmarks for multimodal memory reasoning largely evaluate systems within pre-assembled contexts, but under-evaluate whether agents can use ev

model-releasesarxiv-cs-cl
18 May 2026
Model Releases

Social-Mamba: Socially-Aware Trajectory Forecasting with State-Space Models

DGX agent

arXiv:2605.15424v1 Announce Type: new Abstract: Human trajectory forecasting is crucial for safe navigation in crowded environments, requiring models that balance accuracy with computational efficienc

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

SOLAR: Self-supervised Joint Learning for Symmetric Multimodal Retrieval

DGX agent

arXiv:2605.15868v1 Announce Type: new Abstract: In this work, we address the critical yet underexplored challenge of symmetric multimodal-to-multimodal (MM2MM) retrieval, where queries and contexts ar

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

Sparse ActionGen: Accelerating Diffusion Policy with Real-time Pruning

DGX agent

arXiv:2601.12894v2 Announce Type: replace-cross Abstract: Diffusion Policy has dominated action generation due to its strong capabilities for modeling multi-modal action distributions, but its multi-s

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

STAR: A Stage-attributed Triage and Repair framework for RCA Agents in Microservices

DGX agent

arXiv:2605.15581v1 Announce Type: new Abstract: LLM-based root cause analysis (RCA) agents have recently emerged as a promising paradigm for incident diagnosis in microservice AIOps. However, their re

model-releasesarxiv-cs-ai
18 May 2026
← Previous
1…305306307308309…472
Next →