AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,595 results
Model Releases

Prompting Amazon Nova 2 for content moderation

DGX agent

In this post, you learn how to prompt Amazon Nova 2 Lite for content moderation using structured and free-form approaches, grounded in the MLCommons AILuminate Assessment Standard. The prompting techn

model-releasesaws-ml-blog
18 May 2026
Model Releases
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Quantization Undoes Alignment: Bias Emergence in Compressed LLMs Across Models and Precision Levels

DGX agent

arXiv:2605.15208v1 Announce Type: cross Abstract: Large Language Models are routinely compressed via post-training quantization to reduce inference costs and memory footprint for cloud and edge deploy

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Quantum Feature Pyramid Gating for Seismic Image Segmentation

DGX agent

arXiv:2605.15370v1 Announce Type: cross Abstract: Accurate salt-body delineation is essential for seismic interpretation because salt structures distort wave propagation, complicate velocity-model bui

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

🚀🚀Qwen3.7 Preview lands on Arena ! Here come Qwen3.7-Max-Preview & Qwen3.7-Plus-Preview. Alibaba now #6 lab in Text, #5 in Vision.⚡️⚡️ Can…

DGX agent

🚀🚀Qwen3.7 Preview lands on Arena ! Here come Qwen3.7-Max-Preview & Qwen3.7-Plus-Preview. Alibaba now #6 lab in Text, #5 in Vision.⚡️⚡️ Can't wait to release Qwen3.7 series models!Stay tuned! @arena Qw

model-releasesqwen--x
18 May 2026
Model Releases

RapidUn: Influence-Driven Parameter Reweighting for Efficient Large Language Model Unlearning

DGX agent

arXiv:2512.04457v2 Announce Type: replace Abstract: Removing specific data influence from large language models (LLMs) remains challenging, as retraining is costly and existing approximate unlearning

model-releasesarxiv-cs-cl
18 May 2026
Model Releases

RAR: Retrieving And Ranking Augmented MLLMs for Visual Recognition

DGX agent

arXiv:2403.13805v2 Announce Type: replace-cross Abstract: CLIP (Contrastive Language-Image Pre-training) uses contrastive learning from noise image-text pairs to excel at recognizing a wide array of c

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Reducing the Safety Tax in LLM Safety Alignment with On-Policy Self-Distillation

DGX agent

arXiv:2605.15239v1 Announce Type: new Abstract: Safety alignment often improves robustness to harmful queries at the cost of reasoning ability, a tradeoff known as the safety tax. A common cause is di

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

Registers Matter for Pixel-Space Diffusion Transformers

DGX agent

arXiv:2605.16147v1 Announce Type: new Abstract: Vision Transformers (ViTs) are known to exhibit high-norm patch-token outliers that degrade feature map quality, a problem effectively mitigated by exti

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

Reinforcement learning for adaptive interior point methods in convex quadratic programming

DGX agent

arXiv:2509.07404v2 Announce Type: replace-cross Abstract: Quadratic programming is a workhorse of modern nonlinear optimization, control, and data science. Although regularized methods offer convergen

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

Representation Without Reward: A JEPA Audit for LLM Fine-Tuning

DGX agent

arXiv:2605.15394v1 Announce Type: cross Abstract: Joint-embedding predictive architectures (JEPAs) propose that a model should learn more useful abstractions when trained to predict latent representat

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Retrieval-Augmented Large Language Models for Schema-Constrained Clinical Information Extraction

DGX agent

arXiv:2605.15467v1 Announce Type: cross Abstract: Conversational nurse-patient transcripts contain actionable observations, but converting these transcripts into structured representations at scale re

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

RoadmapBench: Evaluating Long-Horizon Agentic Software Development Across Version Upgrades

DGX agent

arXiv:2605.15846v1 Announce Type: cross Abstract: Coding agents are increasingly deployed in real software development, where a single version iteration requires months of coordinated work across many

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

RTL-BenchMT: Dynamic Maintenance of RTL Generation Benchmark Through Agent-Assisted Analysis and Revision

DGX agent

arXiv:2605.15537v1 Announce Type: new Abstract: This paper introduces RTL-BenchMT, an agentic framework for dynamically maintaining RTL generation benchmarks. Large Language Models (LLMs) assisted aut

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Rule2DRC: Benchmarking LLM Agents for DRC Script Synthesis with Execution-Guided Test Generation

DGX agent

arXiv:2605.15669v1 Announce Type: new Abstract: Manufacturable chip layouts must satisfy thousands of geometry-based design rules, and design rule checking (DRC) enforces them by running executable DR

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

Run Claude Managed Agents with Vercel Sandbox

DGX agent

This article describes how to run Claude's managed agents within Vercel's Sandbox environment, enabling developers to execute AI agent workloads on Vercel's infrastructure. The integration allows user

model-releasesvercel-blog
18 May 2026
Model Releases

Runtime-Orchestrated Second-Order Optimization for Scalable LLM Training

DGX agent

arXiv:2605.16184v1 Announce Type: cross Abstract: Second-order methods offer an attractive path toward more sample-efficient LLM training, but their practical use is often blocked by the systems cost

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

SaaS-Bench: Can Computer-Use Agents Leverage Real-World SaaS to Solve Professional Workflows?

DGX agent

arXiv:2605.15777v1 Announce Type: new Abstract: Computer-Using Agents (CUAs) are rapidly extending large language models (LLMs) beyond text-based reasoning toward action execution in more complex envi

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

SARVLM: A Vision Language Foundation Model for Semantic Understanding in SAR Imagery

DGX agent

arXiv:2510.22665v3 Announce Type: replace-cross Abstract: Synthetic Aperture Radar (SAR) is a critical imaging modality due to its all-weather operational capability. Although recent advances in self-

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

SDOF: Taming the Alignment Tax in Multi-Agent Orchestration with State-Constrained Dispatch

DGX agent

arXiv:2605.15204v1 Announce Type: new Abstract: Multi-agent orchestration frameworks such as LangChain, LangGraph, and CrewAI route tasks through graph-based pipelines but do not enforce the stage con

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Searching on a Budget: HW-NAS with 10 Latency Probes

DGX agent

arXiv:2504.00663v2 Announce Type: replace Abstract: Existing hardware-aware NAS (HW-NAS) methods typically assume access to precise information circa the target device, either via analytical approxima

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

SemanticOpt: Towards LLM-Based Semantic Black-Box Optimization

DGX agent

arXiv:2510.25404v3 Announce Type: replace-cross Abstract: Optimizing an experimental system can be extremely challenging when each experiment is expensive, time-consuming, or difficult to perform. Exi

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

SGR: A Stepwise Reasoning Framework for LLMs with External Subgraph Generation

DGX agent

arXiv:2605.16117v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated strong capabilities across diverse NLP applications, such as translation, text generation, and question a

model-releasesarxiv-cs-cl
18 May 2026
Model Releases

ShopGym: An Integrated Framework for Realistic Simulation and Scalable Benchmarking of E-Commerce Web Agents

DGX agent

arXiv:2605.16116v1 Announce Type: new Abstract: Developing and evaluating e-commerce web agents requires environments that preserve meaningful task structure while enabling controllable, reproducible,

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

SkillSmith: Compiling Agent Skills into Boundary-Guided Runtime Interfaces

DGX agent

arXiv:2605.15215v1 Announce Type: new Abstract: Recently, skills have been widely adopted in large language model (LLM)-based agent systems across various domains. In existing frameworks, skills are t

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

SkyLink: A Large Vision-Language Model Driven Re-ranking Framework for Cross-View UAV geolocalization

DGX agent

arXiv:2603.08063v3 Announce Type: replace Abstract: Cross-view UAV geolocalization is fundamentally a challenging large-scale image retrieval task, aiming to determine the geographic coordinates of Un

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

Skyra: AI-Generated Video Detection via Grounded Artifact Reasoning

DGX agent

arXiv:2512.15693v2 Announce Type: replace Abstract: The misuse of AI-driven video generation technologies has raised serious social concerns, highlighting the urgent need for reliable AI-generated vid

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

SMMBench: A Benchmark for Source-Distributed Multimodal Agent Memory

DGX agent

arXiv:2605.15710v1 Announce Type: new Abstract: Existing benchmarks for multimodal memory reasoning largely evaluate systems within pre-assembled contexts, but under-evaluate whether agents can use ev

model-releasesarxiv-cs-cl
18 May 2026
Model Releases

Social-Mamba: Socially-Aware Trajectory Forecasting with State-Space Models

DGX agent

arXiv:2605.15424v1 Announce Type: new Abstract: Human trajectory forecasting is crucial for safe navigation in crowded environments, requiring models that balance accuracy with computational efficienc

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

SOLAR: Self-supervised Joint Learning for Symmetric Multimodal Retrieval

DGX agent

arXiv:2605.15868v1 Announce Type: new Abstract: In this work, we address the critical yet underexplored challenge of symmetric multimodal-to-multimodal (MM2MM) retrieval, where queries and contexts ar

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

Sparse ActionGen: Accelerating Diffusion Policy with Real-time Pruning

DGX agent

arXiv:2601.12894v2 Announce Type: replace-cross Abstract: Diffusion Policy has dominated action generation due to its strong capabilities for modeling multi-modal action distributions, but its multi-s

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

STAR: A Stage-attributed Triage and Repair framework for RCA Agents in Microservices

DGX agent

arXiv:2605.15581v1 Announce Type: new Abstract: LLM-based root cause analysis (RCA) agents have recently emerged as a promising paradigm for incident diagnosis in microservice AIOps. However, their re

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Steve Bannon and 60+ Trump allies sign a Humans First-led letter urging Trump to mandate government testing and approval of powerful AI models before release (Ashley Gold/Axios)

DGX agent

Ashley Gold / Axios: Steve Bannon and 60+ Trump allies sign a Humans First-led letter urging Trump to mandate government testing and approval of powerful AI models before release — A group of more tha

model-releasestechmeme
18 May 2026
Model Releases

StippleDiffusion: Capacity-Constrained Stippling using Controlled Diffusion

DGX agent

arXiv:2605.15816v1 Announce Type: cross Abstract: Stipple patterns, point sets whose local density tracks a target image, are traditionally produced by per-density iterative optimizers, which are slow

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

Structure Abstraction and Generalization in a Hippocampal-Entorhinal Inspired World Model

DGX agent

arXiv:2605.15733v1 Announce Type: cross Abstract: Humans abstract experiences into structured representations to facilitate pattern inference and knowledge transfer. While the hippocampal-entorhinal (

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Structure-BiEval: A Self-Supervised, Dual-Track Framework for Decoupling Structure and Content in LLM Evaluation for Web Information Systems

DGX agent

arXiv:2601.19923v2 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) evolve into the core of Web-based autonomous agents and complex Web Information Systems, their ability to fait

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

STS: Efficient Sparse Attention with Speculative Token Sparsity

DGX agent

arXiv:2605.15508v1 Announce Type: cross Abstract: The quadratic complexity of attention imposes severe memory and computational bottlenecks on Large Language Model (LLM) inference. This challenge is p

model-releasesarxiv-cs-cl
18 May 2026
Model Releases

SurvivalPFN: Amortizing Survival Prediction via In-Context Bayesian Inference

DGX agent

arXiv:2605.15488v1 Announce Type: new Abstract: Survival analysis provides a powerful statistical framework for modeling time-to-event outcomes in the presence of censoring. However, selecting an appr

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

SynthRender and IRIS: Open-Source Framework and Dataset for Bidirectional Sim-Real Transfer in Industrial Object Perception

DGX agent

arXiv:2602.21141v2 Announce Type: replace Abstract: Object perception is fundamental for tasks such as robotic material handling and quality inspection. However, modern supervised deep-learning models

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

T2T-LA: A Topology-to-Topology LLM Agent for Graph Learning with Neither Feature Access nor Task Knowledge

DGX agent

arXiv:2512.08964v4 Announce Type: replace Abstract: Graph learning aims to convert data into graph representations, which are fundamental to many problems in machine learning for CAD, where circuits,

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

TACO: General Acrobatic Flight Control via Target-and-Command-Oriented Reinforcement Learning

DGX agent

arXiv:2503.01125v4 Announce Type: replace Abstract: Although acrobatic flight control has been studied extensively, one key limitation of the existing methods is that they are usually restricted to sp

model-releasesarxiv-cs-ro
18 May 2026
Model Releases

Tadpole: Autoencoders as Foundation Models for 3D PDEs with Online Learning

DGX agent

arXiv:2605.15284v1 Announce Type: new Abstract: We introduce Tadpole, a novel foundation model for three-dimensional partial differential equations (PDEs) that addresses key challenges in transferabil

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

There are a lot of coding and reasoning benchmarks for AI agents, but not a lot for document understanding - which is a prerequisite for all…

DGX agent

There are a lot of coding and reasoning benchmarks for AI agents, but not a lot for document understanding - which is a prerequisite for all downstream knowledge work. We released ParseBench ~a month

model-releasesjerry-liu--x
18 May 2026
Model Releases

These kids are serial criminals with a callous disregard for life. If they are ever released from jail they will surely harm again. Austin P…

DGX agent

These kids are serial criminals with a callous disregard for life. If they are ever released from jail they will surely harm again. Austin PD, Travis Co. Sheriff Office & Manor PD did their job. Texas

model-releaseselon-musk--x
18 May 2026
Model Releases

💯 this is why I really like Learning mode in Claude Code I personally use this for all my side projects and it keeps me so much sharper, gr…

DGX agent

💯 this is why I really like Learning mode in Claude Code I personally use this for all my side projects and it keeps me so much sharper, great if you want to use Claude Code but still stay hands-on! /

model-releasesboris-cherny--x
18 May 2026
Model Releases

Today in AI Engineering (May 17) • Nous Research ships Hermes Agent v0.14.0: Grok subs, Codex runtime, Windows beta • LangSmith Engine relea…

DGX agent

Today in AI Engineering (May 17) • Nous Research ships Hermes Agent v0.14.0: Grok subs, Codex runtime, Windows beta • LangSmith Engine releases trace issue clustering, drafts PRs and evals from prod t

model-releasesharrison-chase--x
18 May 2026
Model Releases

TokenButler: Token Importance is Predictable

DGX agent

arXiv:2503.07518v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) rely on the Key-Value (KV) Cache to store token history, enabling efficient decoding of tokens. As the KV-Cache g

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Towards Efficient Large Language Reasoning Models via Extreme-Ratio Chain-of-Thought Compression

DGX agent

arXiv:2602.08324v3 Announce Type: replace Abstract: Chain-of-Thought (CoT) reasoning successfully enhances the reasoning capabilities of Large Language Models (LLMs), yet it incurs substantial computa

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

Transformer Scalability Crisis: The First Comprehensive Empirical Analysis of Performance Walls in Modern Language Models

DGX agent

arXiv:2605.15413v1 Announce Type: new Abstract: Despite the remarkable success of transformer architectures in natural language processing, their scalability limitations remain poorly understood throu

model-releasesarxiv-cs-lg
18 May 2026
← Previous
1…305306307308309…471
Next →