AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,595 results
9 Jun 2026

A Survey of Heterogeneous Graph Neural Networks for Cybersecurity Anomaly Detection

Model ReleasesDGX agent

arXiv:2510.26307v3 Announce Type: replace-cross Abstract: Anomaly detection is a critical task in cybersecurity, where identifying insider threats, access violations, and coordinated attacks is essent

A TIL on using http://agentsview.io to calculate token spending with Claude Fable 5 despite that model not yet being included in the AgentsV…

Model ReleasesDGX agent

A TIL on using http://agentsview.io to calculate token spending with Claude Fable 5 despite that model not yet being included in the AgentsView pricing database https://til.simonwillison.net/llms/agen

A Unifying Framework for Concept-Based Representational Similarity


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

arXiv:2606.09653v1 Announce Type: new Abstract: Learned representations across models and modalities often exhibit striking structural similarities, suggesting shared underlying concept decompositions

ABLE: Representing and Mapping LLMs via Attribution-Based Large-model Embedding

Model ReleasesDGX agent

arXiv:2606.07524v1 Announce Type: cross Abstract: The explosive growth of large language models (LLMs) has created a heterogeneous and poorly documented ecosystem, making systematic model comparison i

Activation Steering Induces Emergent Misalignment: A More Comprehensive Evaluation

Model ReleasesDGX agent

arXiv:2606.08682v1 Announce Type: cross Abstract: Activation steering has emerged as a popular inference-time technique for modulating the behavior of large language models (LLMs). By constructing a s

ACTIVE-o3: Empowering MLLMs with Active Perception via Pure Reinforcement Learning

Model ReleasesDGX agent

arXiv:2505.21457v2 Announce Type: replace-cross Abstract: Active vision, also known as active perception, refers to actively selecting where and how to look in order to gather task-relevant informatio

Adaptive directional gradients for parameterised quantum circuits

Model ReleasesDGX agent

arXiv:2606.09734v1 Announce Type: cross Abstract: Training parameterised quantum circuits (PQCs) on quantum hardware is bottlenecked by the measurement cost of gradient estimation, which under the par

AgentCompile: An LLM-Guided Compiler for Direct CUDA Inference

Model ReleasesDGX agent

arXiv:2606.07665v1 Announce Type: cross Abstract: Transformer inference increasingly depends on specialized compiler and runtime support, but real model graphs still require semantic decisions about w

AgroOmni: A Large-Scale Multi-view Agricultural Dataset for Cross-Scale Multimodal Reasoning

Model ReleasesDGX agent

arXiv:2603.14342v2 Announce Type: replace-cross Abstract: Modern agricultural data is sourced from diverse platforms and spans multiple spatial scales, ranging from ground-level close-up photography t

AI Scientists Are Only as Good as Their Evidence: A Stratified Ablation of Proprietary Data and Reasoning Skills in Drug-Asset Valuation

Model ReleasesDGX agent

arXiv:2606.09556v1 Announce Type: new Abstract: AI Scientist agents are often evaluated as if capability were mainly a function of model quality, prompting, or reasoning scaffolds. We test a different

Alcmean's: Unsupervised community detection using local Laplacian, automatic detection of the number of centers

Model ReleasesDGX agent

arXiv:2606.09100v1 Announce Type: cross Abstract: Community detection is a fundamental problem in the analysis of complex networks. It has applications across social, biological, and financial domains

AliyunConsoleAgent: Training Web Agents in Real-World Cloud Environments via Distillation and Reinforcement Learning

Model ReleasesDGX agent

arXiv:2606.09447v1 Announce Type: new Abstract: We present AliyunConsoleAgent, a web agent framework for automated documentation verification in real-world cloud consoles. Major cloud platforms encomp

AlphaOPT: Formulating Optimization Programs with Self-Improving LLM Experience Library

Model ReleasesDGX agent

arXiv:2510.18428v4 Announce Type: replace Abstract: Optimization modeling underlies critical decision-making across industries, yet remains difficult to automate: natural-language problem descriptions

Also, I found that Hermes Agent + Nemotron 3 Ultra is a mighty combo!

Model ReleasesDGX agent

Also, I found that Hermes Agent + Nemotron 3 Ultra is a mighty combo! Excited to launch a new way to upskill with AI agents. This is how we are making it possible for anyone to learn to build with cod

AMN: An Adaptive Multi-Scale Fusion Network with Boundary and Uncertainty Modeling for Nuclei Segmentation

Model ReleasesDGX agent

arXiv:2606.07633v1 Announce Type: cross Abstract: Accurate classification of nuclei subtypes in histopathology images is critical for downstream tasks including tumor grading, immune infiltrate quanti

Anthropic releases its first Mythos-class model Claude Fable

Model ReleasesDGX agent

Anthropic just announced Claude Fable 5, a new AI model it said is the most powerful model it has ever made widely available. According to the company, Fable 5 'shows exceptional performance in softwa

Anthropic says Claude Fable 5 uses conservative safety classifiers that trigger a fallback to Claude Opus 4.8 in <5% of sessions, in areas like cybersecurity (Anthropic)

Model ReleasesDGX agent

Anthropic: Anthropic says Claude Fable 5 uses conservative safety classifiers that trigger a fallback to Claude Opus 4.8 in <5% of sessions, in areas like cybersecurity — Today we're launching Claude

Anthropic says Fable 5 has invisible safeguards that use prompt modification, steering vectors, or PEFT to limit its effectiveness for building frontier LLMs (Matthias Bastian/The Decoder)

Model ReleasesDGX agent

Matthias Bastian / The Decoder: Anthropic says Fable 5 has invisible safeguards that use prompt modification, steering vectors, or PEFT to limit its effectiveness for building frontier LLMs — Key Poin

Anthropic says Fable 5 is available on Pro, Max, Team, and seat-based Enterprise plans through June 22, after which using Fable 5 will require usage credits (Rebecca Bellan/TechCrunch)

Model ReleasesDGX agent

Rebecca Bellan / TechCrunch: Anthropic says Fable 5 is available on Pro, Max, Team, and seat-based Enterprise plans through June 22, after which using Fable 5 will require usage credits — Anthropic is

Anthropic says red team tests of Fable 5 found no universal jailbreaks, and it will keep first- and third-party user traffic on Mythos-class models for 30 days (Derek B. Johnson/CyberScoop)

Model ReleasesDGX agent

Derek B. Johnson / CyberScoop: Anthropic says red team tests of Fable 5 found no universal jailbreaks, and it will keep first- and third-party user traffic on Mythos-class models for 30 days — Claude

Anthropic sets AI performance records with new Mythos 5, Fable 5 frontier models

Model ReleasesDGX agent

Anthropic PBC today introduced Claude Mythos 5 and Claude Fable 5, two large language models that it says outperform the competition across a wide range of benchmarks. The LLMs are derived from the Cl

APEX4: Efficient Pure W4A4 LLM Inference via Intra-SM Compute Rebalancing

Model ReleasesDGX agent

arXiv:2606.08761v1 Announce Type: cross Abstract: W4A4 quantization promises full utilization of INT4 Tensor Cores, yet group dequantization overhead on CUDA Cores has driven existing systems to mixed

Apple says 79% of all iPhones and 86% of iPhones released in the past four years were running iOS 26 as of June 7, vs. 82% and 88% for iOS 18 before WWDC 2025 (Joe Rossignol/MacRumors)

Model ReleasesDGX agent

Joe Rossignol / MacRumors: Apple says 79% of all iPhones and 86% of iPhones released in the past four years were running iOS 26 as of June 7, vs. 82% and 88% for iOS 18 before WWDC 2025 — Apple has sh

Are Reasoning Vision-Language Models Robust to Semantic Visual Distractions?

Model ReleasesDGX agent

arXiv:2606.08894v1 Announce Type: new Abstract: Reasoning Vision-Language Models (VLMs) achieve strong performance on complex multimodal tasks, but reliable real-world application requires handling vi

ArtiFact: A Large-Scale Multi-Modal Cultural Heritage Dataset

Model ReleasesDGX agent

arXiv:2606.09648v1 Announce Type: cross Abstract: Multi-modal data management has emerged as a central research topic in the database community, spanning data integration, semantic query processing, a

Artificial Intelligence for Mathematical Reasoning: An Integrated Survey of Language Models, Neuro-symbolic Systems, and Verified Discovery

Model ReleasesDGX agent

arXiv:2606.08728v1 Announce Type: new Abstract: Mathematical reasoning has long served as a stringent test of machine intelligence; over the past decade, it has moved from a niche problem within NLP t

ATN3D: Density-Aware LiDAR-Radar Early 3D Object Detection Under Extreme Sparsity

Model ReleasesDGX agent

arXiv:2606.09634v1 Announce Type: cross Abstract: 3D object detection is the backbone of perception for automated vehicles (AV) and broader intelligent transportation systems applications. Long-range

Auditable Graph-Guided Root Cause Analysis for Kubernetes Incidents

Model ReleasesDGX agent

arXiv:2606.08590v1 Announce Type: cross Abstract: Kubernetes incidents are diagnosed reliably only when a root-cause system's reported gains come from incident evidence rather than scenario-specific s

Auditing Proprietary Alignment in Large Language Models: A Comparative Framework Without a Ground-Truth Standard

Model ReleasesDGX agent

arXiv:2606.08381v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly released and deployed through opaque development and deployment pipelines, enabling model providers to i

Aujourd'hui, Cohere a annoncé un nouveau partenariat avec le gouvernement du Québec. Le Québec est depuis longtemps l'un des écosystèmes d'I…

Model ReleasesDGX agent

Aujourd'hui, Cohere a annoncé un nouveau partenariat avec le gouvernement du Québec. Le Québec est depuis longtemps l'un des écosystèmes d'IA les plus importants au monde. Nous sommes fiers de collabo

Automatic Extraction of Structured Information from Brain MRI Reports Using an Open-Weight Large Language Model

Model ReleasesDGX agent

arXiv:2606.07721v1 Announce Type: new Abstract: Objectives: Automatic data extraction from free-text radiology reports enables large-scale research, but few studies assessed the performance of large l

AutoMegaKernel: A Statically-Checked Agent Harness for Self-Retargeting Megakernel Synthesis

Model ReleasesDGX agent

arXiv:2606.09682v1 Announce Type: new Abstract: AutoMegaKernel (AMK) compiles a HuggingFace Llama-family model into a single persistent cooperative CUDA kernel that runs the whole forward pass in one

AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs

Model ReleasesDGX agent

arXiv:2606.07643v1 Announce Type: cross Abstract: Recent advances in Omni-Multimodal Large Language Models (Omni-MLLMs) have enabled strong integration of vision, audio, and language. However, their a

AWS debuts AWS FinOps Agent to help customers optimize their cloud spending

Model ReleasesDGX agent

Amazon Web Services Inc. today introduced an artificial intelligence tool designed to help organizations lower their cloud bills. AWS FinOps Agent is initially available in public preview. It’s the la

Bayesian Optimization of a Multi-Product Chemical Reactor Using Composite Models and Partial Physics Knowledge

Model ReleasesDGX agent

arXiv:2606.08611v1 Announce Type: cross Abstract: We study data-driven real-time economic optimization of a multi-product chemical reactor when no reliable first-principles model is available beyond a

Bayesian Selective Latent Inference for Wastewater-First Influenza Monitoring

Model ReleasesDGX agent

arXiv:2606.09433v1 Announce Type: new Abstract: Wastewater influenza surveillance can reveal community circulation before clinical reporting, but wastewater alone is not a fully identifiable proxy for

Benchmark Datasets for Lead-Lag Forecasting on Social Platforms

Model ReleasesDGX agent

arXiv:2511.03877v2 Announce Type: replace Abstract: Social and collaborative platforms emit multivariate time-series traces in which early interactions -- such as views, likes, or downloads -- are fol

Benchmarking Empirical Privacy Protection for Adaptations of Large Language Models

Model ReleasesDGX agent

arXiv:2606.09401v1 Announce Type: new Abstract: Recent work has applied differential privacy (DP) to adapt large language models (LLMs) for sensitive applications, offering theoretical guarantees. How

Benchmarking Open-Ended Multi-Agent Coordination in Language Agents

Model ReleasesDGX agent

arXiv:2606.08340v1 Announce Type: new Abstract: As language models are increasingly deployed as autonomous agents, they must coordinate with others over long horizons in open-ended interactive tasks.

Benchmarking Vision-Language-Action Models on SO-101: Failure and Recovery Analysis

Model ReleasesDGX agent

arXiv:2606.08881v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have demonstrated strong generalization in robotic manipulation, yet existing evaluations are primarily conducted

Beyond Consistency: Preserving Temporal Structure in Zero-Shot Video Editing

Model ReleasesDGX agent

arXiv:2606.08780v1 Announce Type: new Abstract: Existing zero-shot video editing methods rely on pre-trained diffusion models, successfully achieving spatial control and basic temporal consistency but

Beyond English benchmarks: clinical llm evaluation in Brazilian Portuguese

Model ReleasesDGX agent

arXiv:2606.07853v1 Announce Type: cross Abstract: Large Language Models are transforming the support for clinical decision and their application in real scenarios. Yet, most benchmarks are conducted i

Beyond Goodhart's Law: A Dynamic Benchmark for Evaluating Compliance in Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2606.07805v1 Announce Type: new Abstract: The rapid evolution of Large Language Models (LLMs) from passive assistants to autonomous, execution-capable agents has introduced critical operational

Beyond Pass Rate: A Multilingual, Execution-Grounded Evaluation of Open Code LLMs

Model ReleasesDGX agent

arXiv:2606.08840v1 Announce Type: new Abstract: Code generation models are typically compared using compact execution benchmarks and aggregate pass rates, but such summaries obscure how performance va

Beyond Pass/Fail: Using Process Mining to Understand How LLMs Resist (and Fail) Red Team Attacks

Model ReleasesDGX agent

arXiv:2606.07833v1 Announce Type: cross Abstract: Standard AI red teaming evaluations reduce adversarial campaigns to a single binary outcome, attack success rate (ASR), not taking into account the se

Bidirectional Semantic Complementary Tool Retrieval for Remote Sensing Agents

Model ReleasesDGX agent

arXiv:2606.07538v1 Announce Type: cross Abstract: Large language model (LLM)-based agents provide a novel paradigm for the automated processing of remote sensing(RS) data. Their success in complex RS

Bidirectional Small-Granularity Search between Code and Text

Model ReleasesDGX agent

arXiv:2606.07519v1 Announce Type: cross Abstract: We introduce the novel task of bidirectional small-granularity search between code and text, where the queries are small snippets of text or code and

BioVid: Autoregressive Video Generation with Biological Behavior Semantic Comprehension

Model ReleasesDGX agent

arXiv:2606.08674v1 Announce Type: cross Abstract: Existing video generation frameworks treat sequence duration as an externally prescribed parameter -- fixed frame counts or text prompts -- producing

BLUE: Toward Better Language Use in Efficient Vision-Language-Action Models for Autonomous Driving

Model ReleasesDGX agent

arXiv:2606.08684v1 Announce Type: new Abstract: We present BLUE, a minimal method for better language use in vision-language-action (VLA) models for autonomous driving (AD). Through extensive analysis

Bokeh Diffusion: Defocus Blur Control in Text-to-Image Diffusion Models

Model ReleasesDGX agent

arXiv:2503.08434v5 Announce Type: replace-cross Abstract: Recent advances in large-scale text-to-image models have revolutionized creative fields by generating visually captivating outputs from textua

BREAKING: Anthropic just dropped Claude Fable 5—this is Mythos, made safe for public release. It is the best coding model in the world. We'v…

Model ReleasesDGX agent

BREAKING: Anthropic just dropped Claude Fable 5—this is Mythos, made safe for public release. It is the best coding model in the world. We've been testing it internally @every for the last week or so

Breaking the Bubble: Asynchronous Pipeline Parallel Training with Bounded Weight Inconsistency

Model ReleasesDGX agent

arXiv:2606.07881v1 Announce Type: new Abstract: Pipeline parallelism is essential for training large neural networks, but existing schedules trade off throughput, memory, and optimization consistency.

Bridged SBI: Correcting Biased Low-Fidelity Posteriors for Cost-Efficient High-Fidelity Inference

Model ReleasesDGX agent

arXiv:2606.09155v1 Announce Type: new Abstract: Accurate calibration of particle-based simulators is crucial for robotic earthwork simulation, but analytical calibration is challenging due to this tas

BSTabDiff: Block-Subunit Diffusion Priors for High-Dimensional Tabular Data Generation

Model ReleasesDGX agent

arXiv:2606.09257v1 Announce Type: cross Abstract: High-Dimensional Low-Sample Size (HDLSS) tabular domains (e.g., omics) are characterized by n ll m, where n = number of samples, and m = number of fea

btw insane amounts of alpha in telling claude code to 'review my code for issues' on Fable rn while it is not pay per use be prepared to be …

Model ReleasesDGX agent

btw insane amounts of alpha in telling claude code to 'review my code for issues' on Fable rn while it is not pay per use be prepared to be in abject horror that you shipped anything to prod without a

BUDDY: BUdget-Driven DYnamic Depth Routing for Adaptive Large Language Model Inference

Model ReleasesDGX agent

arXiv:2606.09514v1 Announce Type: new Abstract: Large language models (LLMs) incur high inference cost due to their depth and parameter scale. Depth pruning can reduce latency by skipping redundant Tr

but but i thought it was magic?

Model ReleasesDGX agent

but but i thought it was magic? What we learned testing Claude Fable/Mythos 5 on Vending-Bench: > Performance: Makes less money than Opus 4.7 and GPT-5.5 > Alignment: A step back. (Opus 4.8 was better

C3VD-DEFCOL: A Deformable Colonoscopy Dataset with Time-Resolved 3D Ground Truth and Realistic Appearance

Model ReleasesDGX agent

arXiv:2606.07891v1 Announce Type: new Abstract: 3D reconstruction could improve colonoscopy by estimating mucosal coverage and alerting clinicians to missed regions during screening. However, algorith

Calibration of Structured Ignorance Certificates for Diagnosing Unknown Unknowns in Reasoning Models

Model ReleasesDGX agent

arXiv:2606.08571v1 Announce Type: cross Abstract: Large language models frequently fail in a characteristic way: rather than acknowledging ignorance, they produce fluent but incorrect answers to quest

CamoSAM2: SAM2-oriented Prompt Auto-Refinement for Video Camouflaged Object Detection

Model ReleasesDGX agent

arXiv:2504.00375v2 Announce Type: replace Abstract: The Segment Anything Model 2 (SAM2), a prompt-guided video foundation model, has remarkably performed in video object segmentation, drawing signific

← Previous
1…153154155156157…377
Next →