AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

86,993Total entries
1Added by human
86,992Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,477 results
28 May 2026

Explanation Generation for Contradiction Reconciliation with LLMs

ResearchDGX agent

arXiv:2603.22735v2 Announce Type: replace Abstract: Existing NLP work commonly treats contradictions as errors to be resolved by choosing which statements to accept or discard. Yet a key aspect of hum

FactReview: Evidence-Grounded Peer Review with Execution-Based Claim Verification

Model ReleasesDGX agent

arXiv:2604.04074v3 Announce Type: replace Abstract: LLM-based reviewing systems typically take only the manuscript as input, leaving literature and code-based claims hard to verify. We present FactRev

Federated Learning for Multivariate Time Series Anomaly Detection in Industrial Automation

Model ReleasesDGX agent

arXiv:2605.27486v1 Announce Type: new Abstract: Federated learning (FL) has broadened the horizon for multivariate time series anomaly detection (MTSAD). However, benchmarking such anomaly detection m

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Finding Miscompiles for Fun, Not Profit

Model ReleasesDGX agent

This article likely discusses the discovery and analysis of compiler bugs or miscompilation errors in software development tools, exploring how developers identify these issues and their implications

Finding Miscompiles for Fun, Not Profit Or: You don’t need access to Claude Mythos to spend $10,000 in an afternoon https://newsletter.semia…

Model ReleasesDGX agent

This post discusses how compiler bugs or miscompiles can be discovered and exploited without expensive AI systems, illustrating that significant computational costs can be incurred quickly when identi

ForestHG-Trace: Traceable Long-Horizon Ecological Reasoning over Large-Scale Forest Scenes

Model ReleasesDGX agent

arXiv:2605.27590v1 Announce Type: new Abstract: Remote sensing question answering (RS-QA) often requires more than direct semantic prediction, especially in large-scale forest scenes where ecological

From Affect to Complex Behavior: Advancing Multimodal Human-Centered AI at the 10th ABAW Workshop & Competition

SafetyDGX agent

arXiv:2605.27451v1 Announce Type: new Abstract: The 10th Affective & Behavior Analysis in-the-Wild (ABAW) Workshop and Competition, held at CVPR 2026, continues to advance research on modelling, analy

From Detection to Mechanism: Cross-Attention Graph Neural Networks Enable Drug-Drug Interaction Type Prediction An Ablation Study with Acetylsalicylic Acid Validation

Model ReleasesDGX agent

arXiv:2605.27861v1 Announce Type: cross Abstract: Predicting whether two drugs interact (binary detection) is a substantially dif- ferent task from predicting the mechanism type of that interaction (m

From Talking to Singing: A New Challenge for Audio-Visual Deepfake Detection

TutorialsDGX agent

arXiv:2605.27944v1 Announce Type: new Abstract: With rapid advances in audio-visual generative models, reliable forgery detection becomes increasingly critical. Existing methods for audio-visual deepf

FundaPod: A Multi-Persona Agent Pod Platform with Knowledge Graph Memory for AI-Assisted Fundamental Investment Research

AgentsDGX agent

arXiv:2605.27864v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly applied in finance, yet most existing work emphasizes trading signals or financial NLP tasks centered on p

GenSBI: Generative Methods for Simulation-Based Inference in JAX

ResearchDGX agent

arXiv:2605.27499v1 Announce Type: new Abstract: Flow and diffusion generative models have established themselves as widely adopted density estimators for simulation-based inference (SBI), extending na

Global Policy-Space Response Oracles for Two-Player Zero-Sum Games

Model ReleasesDGX agent

arXiv:2605.28273v1 Announce Type: new Abstract: The Policy-Space Response Oracles (PSRO) framework scales equilibrium computation to large zero-sum games by iteratively expanding a restricted strategy

Got a Secret? LLM Agents Can't Keep It: Evaluating Privacy in Multi-Agent Systems

SafetyDGX agent

arXiv:2605.27766v1 Announce Type: new Abstract: LLM safety evaluations predominantly test models in isolation, yet deployed AI agents increasingly operate within persistent social environments alongsi

GRADE: Generalizable Reasoning-Aware Dialogue Evaluation for AI Tutors

ResearchDGX agent

arXiv:2605.27866v1 Announce Type: new Abstract: Evaluating AI tutor responses requires more than factual correctness: tutors must identify mistakes, locate errors, provide guidance, and offer actionab

GradientStabilizer:Fix the Norm, Not the Gradient

Model ReleasesDGX agent

arXiv:2502.17055v4 Announce Type: replace-cross Abstract: Training instability in modern deep learning systems is frequently triggered by rare but extreme gradient-norm spikes, which can induce oversi

Graph Neural Networks for Source Detection: A Review and Benchmark Study

Model ReleasesDGX agent

arXiv:2512.20657v2 Announce Type: replace-cross Abstract: The source detection problem arises when an epidemic process unfolds over a contact network, and the objective is to identify its point of ori

HarmoVid: Relightful Video Portrait Harmonization

Local AiDGX agent

arXiv:2605.28811v1 Announce Type: new Abstract: We present a method for harmonizing the lighting of a foreground video to match a target background scene, adjusting shadows, color tone, and illuminati

hello beloved tasteful users, do you like how much claude thinks on your tasks? would love examples of it thinking too much or too little

Model ReleasesDGX agent

This post asks Claude users for feedback about the length and depth of Claude's reasoning on tasks, soliciting examples where Claude's thinking might be excessive or insufficient. The inquiry aims to

Here Opus 4.8 built and play-tested a new RPG in Claude Code, including 3 PDF manuals and adventures, playtest notes, a website, and a playa…

Model ReleasesDGX agent

Here Opus 4.8 built and play-tested a new RPG in Claude Code, including 3 PDF manuals and adventures, playtest notes, a website, and a playable solo adventure - then put it all on Netlify. No feedback

HO-SFL: Hybrid-Order Split Federated Learning with Backprop-Free Clients and Dimension-Free Aggregation

ResearchDGX agent

arXiv:2603.14773v2 Announce Type: replace-cross Abstract: Fine-tuning large models on edge devices is severely hindered by the memory-intensive backpropagation (BP) in standard frameworks like federat

Holy smokes! Polymarket was not trolling. 500M accidental Claude spend in one month! Scoop from @MadisonMills22 @axios

Model ReleasesDGX agent

Holy smokes! Polymarket was not trolling. 500M accidental Claude spend in one month! Scoop from @MadisonMills22 @axios NEW: AI consultant reveals a client accidentally spent $500,000,000.00 in a singl

How Endava builds an agentic organization with Codex

AgentsDGX agent

Endava, a software services company, leverages OpenAI's Codex to transform its organizational operations into an agentic model where AI agents autonomously handle tasks and decision-making. The approa

how it started how it’s going

Model ReleasesDGX agent

'How it started, how it's going' is a popular internet meme format that compares two contrasting images or states—typically showing an initial hopeful or humble beginning alongside a current outcome t

How the University of Central Oklahoma is using AI to streamline analysis of complex criminal cases

Model ReleasesDGX agent

In the high-stakes world of forensic science, time is the enemy of justice. The University of Central Oklahoma (UCO) Forensic Science Institute (FSI) was looking for an innovative AI solution that cou

https://mistral.ai/news/ai-now-summit-2026/

Model ReleasesDGX agent

Mistral AI announced its participation in or perspective on the AI Now Summit 2026, likely discussing developments in AI safety, ethics, or industry trends relevant to the conference. The announcement

I had Opus 4.8 in Claude Code write a sophisticated, if minor, academic paper from a archive of hundreds of de-identified research files fro…

Model ReleasesDGX agent

I had Opus 4.8 in Claude Code write a sophisticated, if minor, academic paper from a archive of hundreds of de-identified research files from years ago I had to use GPT-5.5 Pro as a reviewer, it spott

I signed up for another SaaS

Model ReleasesDGX agent

Ben Kamarov reflects on his decision to sign up for yet another SaaS product, likely discussing his evaluation criteria, the specific tool's features, or lessons learned about SaaS adoption and tool p

I think you’ll really like Opus 4.8 It’s as smart as its benchmarks show but expresses and utilizes that intelligence in a warm and collabor…

Model ReleasesDGX agent

I think you’ll really like Opus 4.8 It’s as smart as its benchmarks show but expresses and utilizes that intelligence in a warm and collaborative way. Workflows are a great way to utilize it- I’m hook

if you replace billions with millions, this sounds like any other high-growth startup fundraise announcement 😉

Model ReleasesDGX agent

if you replace billions with millions, this sounds like any other high-growth startup fundraise announcement 😉 We've raised 65 billion in Series H funding at a 965 billion post-money valuation, led by

In the last 30 days alone: – Microsoft cancelled most of its Claude Code licenses, citing cost – Uber burned through its entire 2026 AI budg…

Model ReleasesDGX agent

In the last 30 days alone: – Microsoft cancelled most of its Claude Code licenses, citing cost – Uber burned through its entire 2026 AI budget in 4 months – Uber's COO publicly said AI costs are 'hard

Integrating Inductive Biases in Transformers via Distillation for Financial Time Series Forecasting

Local AiDGX agent

arXiv:2603.16985v2 Announce Type: replace Abstract: Transformer-based models have been widely adopted for time-series forecasting due to their high representational capacity and architectural flexibil

Internally Referenced Low-Light Enhancement

Model ReleasesDGX agent

arXiv:2605.28605v1 Announce Type: new Abstract: Self-supervised low-light image enhancement (LLIE) is highly appealing as it eliminates the reliance on external paired data. However, the lack of exter

IRPO: Boosting Image Restoration via Post-training GRPO

ResearchDGX agent

arXiv:2512.00814v3 Announce Type: replace Abstract: Post-training has become effective for high-level generation, but its role in low-level vision remains underexplored. Existing image restoration met

Janus-LoRA: A Balanced Low-Rank Adaptation for Continual Learning

Model ReleasesDGX agent

arXiv:2605.28495v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) has emerged as a promising paradigm for Continual Learning. It independently updates its low-rank factors (A and B), creating

just noticed today - the dataset is already past 1k+ downloads. opensource / openresearch ftw ! @evo__hq would be opensourcing as many datas…

Model ReleasesDGX agent

just noticed today - the dataset is already past 1k+ downloads. opensource / openresearch ftw ! @evo__hq would be opensourcing as many datasets, evals and autoresearch runs as we can in our pursuit of

Knowledge Dependency Estimation for Reliable Question Answering

ResearchDGX agent

arXiv:2605.28047v1 Announce Type: new Abstract: Reliable question answering requires identifying not only whether an answer is correct, but also which available knowledge the prediction depends on. In

Learning Compositional Latent Structure with Vector Networks

Model ReleasesDGX agent

arXiv:2605.28007v1 Announce Type: cross Abstract: Deep networks are powerful function approximators, but they typically store many different computations in shared weight matrices, making it difficult

Learning Deliberately, Acting Intuitively: Unlocking Test-Time Reasoning in Multimodal LLMs

SafetyDGX agent

arXiv:2507.06999v2 Announce Type: replace-cross Abstract: Reasoning is essential for large language models (LLMs), especially in complex tasks such as mathematical problem solving. However, multimodal

Ligand-Conditioned Discrete Diffusion for Protein Sequence-Structure Co-Design

ResearchDGX agent

arXiv:2605.27413v1 Announce Type: cross Abstract: Proteins perform their biological functions through three-dimensional structures encoded by amino acid sequences, and ligand-binding protein co-design

Local Reachy Mini conversations wireless looks like magic! You can bring your friend around the house and get WOW effect from anyone! 🔥 Tha…

Model ReleasesDGX agent

Local Reachy Mini conversations wireless looks like magic! You can bring your friend around the house and get WOW effect from anyone! 🔥 Thanks @andimarafioti for the blog post on how to set this up: h

LUVE : Latent-Cascaded Ultra-High-Resolution Video Generation with Dual Frequency Experts

ResearchDGX agent

arXiv:2602.11564v2 Announce Type: replace Abstract: Recent advances in video diffusion models have significantly improved visual quality, yet ultra-high-resolution (UHR) video generation remains a for

LV-OSD: Language-Vision-Complementary Open-Set Object Detection

Model ReleasesDGX agent

arXiv:2605.28271v1 Announce Type: new Abstract: Object detection is an important task in computer vision, which aims to detect the objects of interest. through the given category list or query images.

markdown-svg-renderer

Model ReleasesDGX agent

Tool: markdown-svg-renderer A slightly customized Markdown rendering tool with special treatment for fenced code SVG blocks - it both renders the image and provides a tab for switching to the code vie

MaskClaw: Edge-Side Personalized Privacy Arbitration for GUI Agents with Behavior-Driven Skill Evolution

Model ReleasesDGX agent

arXiv:2605.28646v1 Announce Type: cross Abstract: GUI agents rely on screenshots to infer intent and operate across applications, but these screenshots often contain private messages, medical records,

MCTS-Judge: Test-Time Scaling in LLM-as-a-Judge for Code Correctness Evaluation

ResearchDGX agent

arXiv:2502.12468v2 Announce Type: replace-cross Abstract: The LLM-as-a-Judge paradigm shows promise for evaluating generative content but lacks reliability in reasoning-intensive scenarios, such as pr

MemCog: From Memory-as-Tool to Memory-as-Cognition in Conversational Agents

Model ReleasesDGX agent

arXiv:2605.28046v1 Announce Type: new Abstract: Existing agent memory systems universally follow what we term a Memory-as-Tool paradigm where a single query triggers one-shot retrieval of flat passage

MeniOmni: A Structured Multimodal Benchmark for Holistic Meniscus Injury Assessment

Model ReleasesDGX agent

arXiv:2605.28161v1 Announce Type: new Abstract: Clinical diagnosis of meniscus injuries requires radiologists to integrate volumetric MRI evidence with patient context (e.g., sex, age, BMI) and to pro

Mistral says it is accelerating superintelligence development to ensure Europe's independence from US tech giants, and signs deals to supply Airbus and BMW (Sam Schechner/Wall Street Journal)

Model ReleasesDGX agent

Sam Schechner / Wall Street Journal: Mistral says it is accelerating superintelligence development to ensure Europe's independence from US tech giants, and signs deals to supply Airbus and BMW — Frenc

Mitigating Cross-Lingual Cultural Inconsistencies in LLMs via Consensus-Driven Preference Optimisation

SafetyDGX agent

arXiv:2605.12515v2 Announce Type: replace Abstract: Despite their impressive capabilities, multilingual large language models (MLLMs) frequently exhibit inconsistent behaviour when the prompt's langua

Narrative Flattening: How Post-Training Compresses Thematic, Affective, and Stylistic Variation in LLM Fiction

SafetyDGX agent

arXiv:2605.27878v1 Announce Type: new Abstract: Large language models produce fluent fiction, yet their creative output is widely seen as flat. We ask where this quality originates in the training and

New in Claude Code (research preview): dynamic workflows. Claude writes an orchestration script on the fly, then spins up a large fleet of c…

Model ReleasesDGX agent

New in Claude Code (research preview): dynamic workflows. Claude writes an orchestration script on the fly, then spins up a large fleet of coordinated subagents in parallel to take on your most comple

On the Intrinsic Limits of Transformer Image Embeddings in Non-Solvable Spatial Reasoning

Model ReleasesDGX agent

arXiv:2601.03048v2 Announce Type: replace-cross Abstract: Vision Transformers (ViTs) excel in semantic recognition but exhibit systematic failures in spatial reasoning tasks such as mental rotation. W

On the Subgaussianity of Quantized Linear Maps: An AI-Assisted Note

Model ReleasesDGX agent

arXiv:2605.27563v1 Announce Type: cross Abstract: This short note presents a dimension-independent subgaussian concentration bound for Gaussian vectors under coordinate-wise nonlinear mappings. Discov

Optimal LTLf Synthesis

Model ReleasesDGX agent

arXiv:2605.11544v2 Announce Type: replace Abstract: Strategy synthesis typically follows an all-or-nothing paradigm, returning unrealisable whenever a specification cannot be guaranteed in an uncertai

Optimal ridge regularization revisited

Model ReleasesDGX agent

arXiv:2605.28679v1 Announce Type: new Abstract: We consider L^2-regularized linear (ridge) regression over a finite data sample X with bounded covariance and linear prediction targets y with additive

Opus 4.8 formulated the hypotheses in advance, conducting data cleaning, did research on references, conducted analyses, did robustness chec…

Model ReleasesDGX agent

Opus 4.8 formulated the hypotheses in advance, conducting data cleaning, did research on references, conducted analyses, did robustness checks, and put out the whole paper in LaTEX style. GPT-5.5 foun

PEAM: Parametric Embodied Agent Memory through Contrastive Internalization of Experience in Minecraft

Model ReleasesDGX agent

arXiv:2605.27762v1 Announce Type: new Abstract: We present PEAM, a Parametric Embodied Agent Memory framework in Minecraft that transforms agent memory from inference-time retrieval into parameter-res

Personal Visual Memory from Explicit and Implicit Evidence

Model ReleasesDGX agent

arXiv:2605.28806v1 Announce Type: cross Abstract: Long-term memory is increasingly important for personalized AI agents, yet existing benchmarks and methods remain largely text-centric. Even when imag

Personalized Observation Normalization for Federated Reinforcement Learning in Simulation Environments with Heterogeneity

Model ReleasesDGX agent

arXiv:2605.27385v1 Announce Type: cross Abstract: Federated reinforcement learning (FedRL) enables multiple agents to collaboratively train a global policy without sharing raw data, making it ideal fo

POINav: Benchmarking and Enhancing Final-Meters Arrival in Real-World Vision-Language Navigation

Model ReleasesDGX agent

arXiv:2605.28237v1 Announce Type: cross Abstract: Real-world navigation is fundamentally driven by Points of Interest (POIs), yet reaching a precise POI remains a critical 'final-meters' challenge. Ex

← Previous
1…661662663664665…1042
Next →