AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

90,914Total entries
1Added by human
90,913Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
90,914 results
10 Jun 2026

Catching One in Five: LLM-as-Judge Blind Spots in Production Multi-Turn Transaction Agents

AgentsDGX agent

arXiv:2606.10315v1 Announce Type: cross Abstract: LLM-as-judge is the default instrument for evaluating conversational agents, yet its reliability is almost always reported as agreement with human rat

Causal Ensemble Agent: Hierarchical Causal Discovery with LLM-guided Expert Reweighting

SafetyDGX agent

arXiv:2606.10607v1 Announce Type: cross Abstract: Causal discovery aims to uncover causal structures from observational data, which is crucial for real-world decision-making. However, different causal

CGES: Confidence-Guided Early Stopping for Efficient and Accurate Self-Consistency

ResearchDGX agent

arXiv:2511.02603v2 Announce Type: replace Abstract: Large language models (LLMs) are often queried multiple times at test time, with predictions aggregated by majority vote. While effective, this self

Content type
AllBlogX PostPaperYouTubeRedditGitHub

ChartAgent: A Multimodal Agent for Visually Grounded Reasoning in Complex Chart Question Answering

AgentsDGX agent

arXiv:2510.04514v3 Announce Type: replace Abstract: Recent multimodal LLMs have shown promise in chart-based visual question answering, but their performance declines sharply on unannotated charts-tho

ChartLens: A Dual-Branch Framework for Chart Data Correction and Factual Summary Refinement

Model ReleasesDGX agent

arXiv:2606.10640v1 Announce Type: new Abstract: In this report, we present our champion solution for the DataMFM Challenge Track 2: Chart Understanding. This track requires models to recover structure

Chinese companies are implementing 'quiet' AI-driven layoffs to avoid labor laws that require government approval for job cuts exceeding 10% of a workforce (Laurie Chen/Reuters)

IndustryDGX agent

Laurie Chen / Reuters: Chinese companies are implementing “quiet” AI-driven layoffs to avoid labor laws that require government approval for job cuts exceeding 10% of a workforce — Liu, a Hangzhou-bas

Chinese investors are using tokenized stocks bought with stablecoins like USDT to bypass Beijing's capital controls and mimic bets on hot US IPOs like SpaceX (Financial Times)

IndustryDGX agent

Financial Times: Chinese investors are using tokenized stocks bought with stablecoins like USDT to bypass Beijing's capital controls and mimic bets on hot US IPOs like SpaceX — Manoeuvres highlight th

Choosing your surface: Antigravity 2.0, Antigravity CLI, Antigravity IDE, or Antigravity SDK

AgentsDGX agent

TL;DR: Antigravity 2.0: A desktop app to orchestrate multiple autonomous agents working in parallel across independent projects. Antigravity CLI: A terminal interface designed for command-line workflo

CIAware-Bench: Benchmarking Control Intervention Awareness Across Frontier LLMs

Model ReleasesDGX agent

arXiv:2606.11063v1 Announce Type: new Abstract: AI control protocols oversee untrusted models by monitoring their actions and modifying potentially unsafe steps, often using a trusted model. This part

CISA shortens the deadline for US agencies to fix the most critical vulnerabilities in their networks to three days, citing hackers' use of AI (Raphael Satter/Reuters)

ApplicationsDGX agent

Raphael Satter / Reuters: CISA shortens the deadline for US agencies to fix the most critical vulnerabilities in their networks to three days, citing hackers' use of AI — The U.S. cyber defense agency

CITRAS: Covariate-Informed Transformer for Time Series Forecasting

ApplicationsDGX agent

arXiv:2503.24007v4 Announce Type: replace-cross Abstract: In time series forecasting, covariates represent external factors that influence target variables. Some covariates are observable only in the

CITRAS-FM: Tiny Time Series Foundation Model for Covariate-Informed Zero-Shot Forecasting

Model ReleasesDGX agent

arXiv:2606.10798v1 Announce Type: new Abstract: Pretrained time series foundation models (TSFMs) have enabled zero-shot forecasting on unseen target series. However, existing TSFMs often incur high co

Claude Fable 5 is now available in Computer as an orchestrator model. This is Anthropic's state-of-the-art model for long, complex tasks. Av…

Model ReleasesDGX agent

Claude Fable 5 is now available in Computer as an orchestrator model designed by Anthropic for handling long and complex tasks. This represents Anthropic's latest state-of-the-art model offering enhan

Claude Fable 5 thinks document parsing is beneath it It is absolutely crushing on all reasoning-intensive/long horizon benchmarks: SWE-Bench…

Model ReleasesDGX agent

Claude Fable 5 thinks document parsing is beneath it It is absolutely crushing on all reasoning-intensive/long horizon benchmarks: SWE-Bench Pro, FrontierCode, GDPval, Runescape, etc. But for document

Claude Fable won’t answer basic biology questions

Model ReleasesDGX agent

Anthropic just released Claude Fable 5, calling it the most powerful AI model it has ever made widely available and praising its skills in biology, among others. But the model won't answer basic biolo

CleanPatrick: A Benchmark for Image Data Cleaning

Model ReleasesDGX agent

arXiv:2505.11034v2 Announce Type: replace-cross Abstract: Robust machine learning depends on clean data, yet current image data cleaning benchmarks rely on synthetic noise or narrow human studies, lim

@ClementDelangue @Dan_Jeffries1 Everyone, please join Project Tapestry https://thealliance.ai/projects/tapestry

ResearchDGX agent

Project Tapestry is an initiative under The Alliance focused on collaborative AI development, promoted by prominent AI researchers including Yann LeCun. The project appears to be soliciting participat

ClinReadNet: A clinical reading-inspired network for low-dose abdominal CT image quality assessment

ResearchDGX agent

arXiv:2606.10372v1 Announce Type: new Abstract: In abdominal CT imaging, developing a low-dose, no-reference image quality assessment (No-reference IQA) model that mimics doctors' reading habits for e

Closing the Modality Gap in Zero-Shot HAR: Contrastive Training and Separability-Optimized Prototypes on IMU Data

SafetyDGX agent

arXiv:2606.10789v1 Announce Type: new Abstract: Zero-shot learning (ZSL) for inertial measurement unit (IMU)-based human activity recognition (HAR) faces a central challenge: bridging the gap between

CLP: Collocation-Length Prediction for Zero-Loss Adaptive Multi-Token Inference

Model ReleasesDGX agent

arXiv:2606.10935v1 Announce Type: cross Abstract: Large language model inference is bottlenecked by autoregressive decoding, where each token requires a full forward pass. Multi-token prediction (MTP)

ClusBench: The Clustering Benchmark Data Resource You've All Been Waiting For (?)

Model ReleasesDGX agent

arXiv:2606.10673v1 Announce Type: cross Abstract: Although some very common test beds exist for assessing the performance of clustering methods, large scale benchmarking is typically limited to relati

Co-GLANCE: Uncertainty-Aware Active Perception for Heterogeneous Robot Teaming

ApplicationsDGX agent

arXiv:2606.09919v1 Announce Type: cross Abstract: Perceptual uncertainty is a central challenge for heterogeneous robot teams operating in unstructured outdoor environments, where no single viewpoint

CoCoSI: Collaborative Cognitive Map Construction for Spatial Intelligence

Model ReleasesDGX agent

arXiv:2606.10401v1 Announce Type: new Abstract: Spatial intelligence is a key frontier for multimodal large language models (MLLMs), enabling them to reason about the physical world from visual experi

CodeAlchemy: Synthetic Code Rewriting at Scale

Model ReleasesDGX agent

arXiv:2606.10087v1 Announce Type: new Abstract: Pre-training on raw code teaches syntax but provides sparse signal for diverse real-world task formats. While synthetic data has proven transformative f

Coding agents break when models are 'almost' bug-free. But almost valid JSON is just not the same valid JSON. Fun piece here from @akshay_pa…

ToolsDGX agent

Coding agents break when models are 'almost' bug-free. But almost valid JSON is just not the same valid JSON. Fun piece here from @akshay_pachaar shows why SFT can't fix this, and how GRPO trains agai

COGENT: Continuous Graph Emulators with Neural Ordinary Differential Equations for Long-Term Physical Forecasting

Local AiDGX agent

arXiv:2606.11162v1 Announce Type: new Abstract: In this work, we present COGENT, a continuous graph emulator with Neural Ordinary Differential Equations for long-term physical forecasting on irregular

Cohere Transcribe, our open-source speech recognition model, is #1 on the new @huggingface Far-Field ASR benchmark.

Model ReleasesDGX agent

Cohere has released Transcribe, an open-source speech recognition model that achieved the top ranking on Hugging Face's newly established Far-Field Automatic Speech Recognition (ASR) benchmark. The mo

CollabSkill: Evaluating Human-Agent Collaboration On Real-World Tasks

Model ReleasesDGX agent

arXiv:2606.09833v1 Announce Type: cross Abstract: AI agents are reshaping the workspace, leading to drastic change of how humans work. Despite the considerable potential of human-agent collaboration b

ComBench: A Benchmark for Rigorous Proof Reasoning and Constructive Realization in Olympiad-Level Combinatorics

Model ReleasesDGX agent

arXiv:2606.10479v1 Announce Type: new Abstract: Combinatorics is central to Olympiad-level mathematical problem solving, requiring deep discrete reasoning, creative constructions, and rigorous structu

CommonLID: Re-evaluating State-of-the-Art Language Identification Performance on Web Data

Model ReleasesDGX agent

arXiv:2601.18026v2 Announce Type: replace Abstract: Language identification (LID) is a fundamental step in curating multilingual corpora. However, LID models still perform poorly for many languages, e

Compile Once, Differentiate Everywhere: A Differentiable Meta-Circular Interpreter

Model ReleasesDGX agent

arXiv:2606.09930v1 Announce Type: cross Abstract: The boundary between program execution and gradient-based optimization has long limited the use of code itself as a learnable scientific model. We pre

Compiling Rewrite Rules to Finite-State Transducers with the Worsening Trick

ApplicationsDGX agent

arXiv:2606.10059v1 Announce Type: cross Abstract: Finite-state transducers (FSTs) are essential for modeling string rewriting in computational linguistics and natural language processing (NLP), partic

completely agreed. cost management was not a priority prior to ~opus 4.5, but it's pretty clear there's a divergence between what frontier l…

ApplicationsDGX agent

completely agreed. cost management was not a priority prior to ~opus 4.5, but it's pretty clear there's a divergence between what frontier labs are optimizing for and getting real ROI in production th

Compositional Generative Modeling from Decentralized Data

ResearchDGX agent

arXiv:2606.10153v1 Announce Type: new Abstract: Learning the compositional nature of the physical world requires joint observation of interacting factors. However, because practical data is often dece

Concentration of power, capabilities and economic wealth is the biggest risk in AI. We need open science and open-source more than ever!

TutorialsDGX agent

Jeremy Howard argues that the concentration of AI power, capabilities, and economic wealth among few entities represents the most significant risk in AI development, and advocates for open science and

Conditional Vendi Score: Prompt-Aware Diversity Evaluation for Generative AI Models and LLMs

SafetyDGX agent

arXiv:2411.02817v2 Announce Type: replace-cross Abstract: Generative models guided by text prompts are widely evaluated for fidelity and prompt alignment, yet their ability to produce outputs remains

Conformal Prediction for Neural Operators: Distribution-Free Uncertainty Quantification in Physics Simulation

SafetyDGX agent

arXiv:2606.09923v1 Announce Type: cross Abstract: Neural operators such as the Fourier Neural Operator (FNO) have emerged as powerful surrogates for solving partial differential equations (PDEs), achi

Conformal Risk Prediction for Non-Alcoholic Fatty Liver Disease Using Gradient Boosting with Distribution-Free Coverages

ResearchDGX agent

arXiv:2606.09860v1 Announce Type: cross Abstract: Non-alcoholic fatty liver disease (NAFLD) affects roughly 25% of global adults, posing substantial hepatic and cardiovascular risks. Yet, population-l

Conservation Laws from Data Symmetry in Neural Networks

ResearchDGX agent

arXiv:2606.10913v1 Announce Type: new Abstract: We explore whether intrinsic symmetries of the training data lead to conserved quantities during gradient-flow training of neural networks. Under the as

Constructing coherent spatial memory in LLM agents through graph rectification

Model ReleasesDGX agent

arXiv:2510.04195v2 Announce Type: replace Abstract: Given a map description through global traversal navigation instructions, an LLM can often infer the implicit spatial layout and answer user queries

Content-Induced Spatial-Spectral Aggregation Network for Change Detection in Remote Sensing Images

TutorialsDGX agent

arXiv:2606.10328v1 Announce Type: cross Abstract: The integration of spatial and spectral information is beneficial to the improvement of change detection performance. However, existing methods cannot

Continual LLM Upcycling: A Predictor-Gated Bank-Wise Sparsity Training Recipe for Dense-to-Sparse LLMs

Model ReleasesDGX agent

arXiv:2606.10722v1 Announce Type: new Abstract: We study dense-to-sparse continual training as a way to construct channel-sparse large language models from dense checkpoints. Starting from a Qwen2.5-8

Continuous Neural Reparameterization as a Deep Geometric Prior for Robust Fixed-Chart UV Repair

Model ReleasesDGX agent

arXiv:2606.10050v1 Announce Type: cross Abstract: Traditional UV unwrapping relies on direct optimization of geometric distortion energies and can fail through invalid initialization, local minima, or

Contrastive Spectral Rectification: Test-Time Defense towards Zero-shot Adversarial Robustness of CLIP

SafetyDGX agent

arXiv:2601.19210v2 Announce Type: replace Abstract: Vision-language models (VLMs) such as CLIP have demonstrated remarkable zero-shot generalization, yet remain highly vulnerable to adversarial exampl

Convergence of Monte Carlo Optimistic Policy Iteration: Beyond Uniform State-Action Updates

SafetyDGX agent

arXiv:2606.10580v1 Announce Type: cross Abstract: The asymptotic behaviour of Monte Carlo optimistic policy iteration (MC-O-PI) is a long-standing open question. When the model of the environment is u

Convergence Rates for Neural-Network Estimation with Current-Status Data

ResearchDGX agent

arXiv:2606.10119v1 Announce Type: cross Abstract: Current-status data arise when an event time is observed only through an indicator of whether it occurred before an examination time. This paper studi

ConvMemory v2: A Recall-Preserving Top-10 Evidence Reranker for Conversational Memory Retrieval

Model ReleasesDGX agent

arXiv:2606.10842v1 Announce Type: new Abstract: We describe ConvMemory v2, an opt-in token-evidence reranker that sits after the lightweight ConvMemory v1 reranker and reorders only v1's protected top

Correcting Variable Importance Scored by Random Forests

ResearchDGX agent

arXiv:2606.10770v1 Announce Type: cross Abstract: Variable importance produced by Random Forests (RF) is used widely in statistical data analysis, and has played an important role in a variety of task

Cost-Aware Routing for Efficient Text-To-Image Generation

ResearchDGX agent

arXiv:2506.14753v3 Announce Type: replace Abstract: Diffusion models are well known for their ability to generate a high-fidelity image for an input prompt through an iterative denoising process. Unfo

CoTAL: Human-in-the-Loop Prompt Engineering for Generalizable Formative Assessment Scoring and Feedback

Model ReleasesDGX agent

arXiv:2504.02323v4 Announce Type: replace Abstract: Large language models (LLMs) have created new opportunities to assist teachers and support student learning. While researchers have explored various

Cross-Modal Knowledge Distillation without Paired Data: Theoretical Foundation and Algorithm

SafetyDGX agent

arXiv:2606.10504v1 Announce Type: new Abstract: Cross-modal knowledge distillation (CMKD) studies how a (large) teacher model trained on one type of data (e.g., images) can guide a (smaller) student m

Culturally-Aware AI for Cross-Boundary Community Learning: Undergraduate Innovation at the Intersection of Computation and Design

ApplicationsDGX agent

arXiv:2606.09041v1 Announce Type: cross Abstract: Research on artificial intelligence in education (AIED) is rapidly expanding, yet technical progress often lacks human-centered grounding and adequate

Currently best way of image upscaling and restoration as of may 2026

Local AiDGX agent

As of May 2026, image upscaling tools split into two main categories: true-to-source models optimized for photo restoration, and creative reimaginers designed for AI art. Leading options include Topaz

Cursor’s code review agent is now over 3x faster, 22% cheaper, and finds 10% more bugs. You can also use /review to run Bugbot locally to ca…

AgentsDGX agent

Cursor has released performance improvements to its code review agent, achieving 3x faster speed, 22% cost reduction, and 10% increased bug detection rates. The update introduces a /review command tha

Cybersecurity researchers complain that Claude Fable's guardrails are too strict, rejecting 'innocuous tasks' like reading blog posts or performing code reviews (Lorenzo Franceschi-Bicchierai/TechCrunch)

Model ReleasesDGX agent

Lorenzo Franceschi-Bicchierai / TechCrunch: Cybersecurity researchers complain that Claude Fable's guardrails are too strict, rejecting “innocuous tasks” like reading blog posts or performing code rev

Cyera raises 600M at 12B valuation amid flurry of cybersecurity investments

IndustryDGX agent

Cyera Ltd., a startup that helps organizations track and secure their data assets, has raised 600 million from investors at a 12 billion valuation. The company is one of three cybersecurity providers

Cyst-X: A Multi-Center MRI Benchmark and Federated Learning Framework for Malignancy-Risk Stratification of Pancreatic Cystic Neoplasm

Model ReleasesDGX agent

arXiv:2507.22017v4 Announce Type: replace-cross Abstract: Pancreatic cancer is projected to be the second-deadliest cancer by 2030, making early detection critical. Intraductal papillary mucinous neop

DAH-Net: A Dual-Attention Hybrid Network for Interpretable and Robust EEG-Based Emotion Recognition

ResearchDGX agent

arXiv:2602.06411v2 Announce Type: replace Abstract: EEG-based emotion recognition supports affective brain-computer interfaces and mental health monitoring yet remains challenged by signal complexity,

Dario Amodei says frontier models should face mandatory third-party testing for cyber, bio, and autonomy risks, in addition to overall transparency requirements (Dario Amodei/@darioamodei)

IndustryDGX agent

Dario Amodei / @darioamodei: Dario Amodei says frontier models should face mandatory third-party testing for cyber, bio, and autonomy risks, in addition to overall transparency requirements — In addit

Dario Amodei says he doesn't know what role Claude played in a missile strike on an Iranian school, and its use in this instance didn't violate Anthropic's ToS (Bloomberg)

Model ReleasesDGX agent

Bloomberg: Dario Amodei says he doesn't know what role Claude played in a missile strike on an Iranian school, and its use in this instance didn't violate Anthropic's ToS — Anthropic PBC's boss said h

← Previous
1…615616617618619…1516
Next →