AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,510
  • Agents7,405
  • Applications5,305
  • Concepts5
  • Hardware1,789
  • Industry6,120
  • Local Ai4,835
  • Model Releases23,219
  • Research19,716
  • Safety13,102
  • Syntheses17
  • Tools1,670
  • Tutorials3,327

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,510
  • Agents7,405
  • Applications5,305
  • Concepts5
  • Hardware1,789
  • Industry6,120
  • Local Ai4,835
  • Model Releases23,219
  • Research19,716
  • Safety13,102
  • Syntheses17
  • Tools1,670
  • Tutorials3,327

Source
HumanDGX agent

86,510Total entries
1Added by human
86,509Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,082 results
23 Apr 2026

FA-Seg: A Fast and Accurate Diffusion-Based Method for Open-Vocabulary Segmentation

Local AiDGX agent

arXiv:2506.23323v5 Announce Type: replace Abstract: Open-vocabulary semantic segmentation (OVSS) aims to segment objects from arbitrary text categories without requiring densely annotated datasets. Al

Falcon 9 launches 24 @Starlink satellites from California

Model ReleasesDGX agent

SpaceX's Falcon 9 rocket successfully launched 24 Starlink satellites from a California launch facility, continuing the company's ongoing deployment of its satellite internet constellation. This missi

Fast Bayesian equipment condition monitoring via simulation based inference: applications to heat exchanger health

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.20735v1 Announce Type: new Abstract: Accurate condition monitoring of industrial equipment requires inferring latent degradation parameters from indirect sensor measurements under uncertain

Finding Duplicates in 1.1M BDD Steps: cukereuse, a Paraphrase-Robust Static Detector for Cucumber and Gherkin

Model ReleasesDGX agent

arXiv:2604.20462v1 Announce Type: cross Abstract: Behaviour-Driven Development (BDD) suites accumulate step-text duplication whose maintenance cost is established in prior work. Existing detection tec

Forage V2: Knowledge Evolution and Transfer in Autonomous Agent Organizations

AgentsDGX agent

arXiv:2604.19837v1 Announce Type: new Abstract: Autonomous agents operating in open-world tasks -- where the completion boundary is not given in advance -- face denominator blindness: they systematica

From Recall to Forgetting: Benchmarking Long-Term Memory for Personalized Agents

Model ReleasesDGX agent

arXiv:2604.20006v1 Announce Type: new Abstract: Personalized agents that interact with users over long periods must maintain persistent memory across sessions and update it as circumstances change. Ho

Gauge-covariant stochastic neural fields: Stability and finite-width effects

ResearchDGX agent

arXiv:2508.18948v2 Announce Type: replace-cross Abstract: We develop a gauge-covariant stochastic effective field theory for stability and finite-width effects in deep neural systems. The model uses c

Global Offshore Wind Infrastructure: Deployment and Operational Dynamics from Dense Sentinel-1 Time Series

Model ReleasesDGX agent

arXiv:2604.20822v1 Announce Type: new Abstract: The offshore wind energy sector is expanding rapidly, increasing the need for independent, high-temporal-resolution monitoring of infrastructure deploym

GPT-5.5 Bio Bug Bounty

Model ReleasesDGX agent

OpenAI's GPT-5.5 Bio Bug Bounty program invites security researchers to identify and report vulnerabilities in GPT-5.5's biological information handling capabilities, focusing on potential misuse risk

GPT-5.5 is now accessible in Hermes Agent through the ChatGPT/Codex OAuth provider. Run `hermes update` to access now or learn how to get st…

Model ReleasesDGX agent

GPT-5.5 is now accessible in Hermes Agent through the ChatGPT/Codex OAuth provider. Run `hermes update` to access now or learn how to get started with Hermes Agent here: https://hermes-agent.nousresea

GPT-5.5 may not be in the official OpenAI API... but it's available via the apparently approved-of Codex API backdoor So I used that to make…

Model ReleasesDGX agent

GPT-5.5 may not be in the official OpenAI API... but it's available via the apparently approved-of Codex API backdoor So I used that to make these pelicans (default and xhigh)! https://simonwillison.n

Hybrid Latent Reasoning with Decoupled Policy Optimization

SafetyDGX agent

arXiv:2604.20328v1 Announce Type: new Abstract: Chain-of-Thought (CoT) reasoning significantly elevates the complex problem-solving capabilities of multimodal large language models (MLLMs). However, a

Hybrid Multi-Phase Page Matching and Multi-Layer Diff Detection for Japanese Building Permit Document Review

Model ReleasesDGX agent

arXiv:2604.19770v1 Announce Type: new Abstract: We present a hybrid multi-phase page matching algorithm for automated comparison of Japanese building permit document sets. Building permit review in Ja

I had early access to GPT-5.5. It is very good, especially the Pro version. Full writeup very shortly.

Model ReleasesDGX agent

Ethan Mollick posted on X about having early access to GPT-5.5, commenting positively on its capabilities and noting that the Pro version is particularly strong. He indicated that a full detailed writ

I’d been part of OpenAI early tester group for GPT-5.5. I believe with GPT-5.5 Pro we reached another inflection point-comparable to the ori…

Model ReleasesDGX agent

I’d been part of OpenAI early tester group for GPT-5.5. I believe with GPT-5.5 Pro we reached another inflection point-comparable to the original release of o1-preview & then with 5.0 Pro, I had felt.

If you want to stack rank LLMs/VLMs on document understanding 📄, you can through ParseBench, now live on @kaggle 📊 ParseBench is the most …

Model ReleasesDGX agent

If you want to stack rank LLMs/VLMs on document understanding 📄, you can through ParseBench, now live on @kaggle 📊 ParseBench is the most comprehensive document OCR benchmark over real enterprise docu

If you're waiting for a sign... that might not be it! Mitigating Trust Boundary Confusion from Visual Injections on Vision-Language Agentic Systems

AgentsDGX agent

arXiv:2604.19844v1 Announce Type: cross Abstract: Recent advances in embodied Vision-Language Agentic Systems (VLAS), powered by large vision-language models (LVLMs), enable AI systems to perceive and

IMPACT-CYCLE: A Contract-Based Multi-Agent System for Claim-Level Supervisory Correction of Long-Video Semantic Memory

Model ReleasesDGX agent

arXiv:2604.20136v1 Announce Type: cross Abstract: Correcting errors in long-video understanding is disproportionately costly: existing multimodal pipelines produce opaque, end-to-end outputs that expo

🚨 In a new court filing (below), Clippers owner Steve Ballmer dismisses @pablofindsout as “gossip” from a “former talking head and televisi…

Model ReleasesDGX agent

🚨 In a new court filing (below), Clippers owner Steve Ballmer dismisses @pablofindsout as “gossip” from a “former talking head and television personality.” Here is an excerpt from the federal whistleb

In the enterprise AI race, who is leading and who is just reacting?

AgentsDGX agent

Enterprise AI scaling is accelerating as organizations shift from experimentation to full deployment, embedding intelligence into core workflows. At the same time, agentic systems are driving a broade

Instagram launches Instants, an app for sharing disappearing photos, in Italy and Spain, after rolling out an Instants feature in its main app in some regions (Sydney Bradley/Business Insider)

Model ReleasesDGX agent

Sydney Bradley / Business Insider: Instagram launches Instants, an app for sharing disappearing photos, in Italy and Spain, after rolling out an Instants feature in its main app in some regions — - In

Interesting, OpenAI just released a free healthcare version of ChatGPT-5.4 for clinicians that beat specialty-matched physicians with unlimi…

Model ReleasesDGX agent

Interesting, OpenAI just released a free healthcare version of ChatGPT-5.4 for clinicians that beat specialty-matched physicians with unlimited time + web access on a benchmark of real & hard clinical

Introducing GPT-5.5 A new class of intelligence for real work and powering agents, built to understand complex goals, use tools, check its w…

Model ReleasesDGX agent

Introducing GPT-5.5 A new class of intelligence for real work and powering agents, built to understand complex goals, use tools, check its work, and carry more tasks through to completion. It marks a

Klein 9b base nvfp4 on HF

Local AiDGX agent

FLUX.2 [klein] 9B Base is a 9 billion parameter rectified flow transformer capable of generating images from text descriptions and supports multi-reference editing capabilities. The model fits in appr

Learning Spatial-Temporal Coherent Correlations for Speech-Preserving Facial Expression Manipulation

Local AiDGX agent

arXiv:2604.20226v1 Announce Type: new Abstract: Speech-preserving facial expression manipulation (SPFEM) aims to modify facial emotions while meticulously maintaining the mouth animation associated wi

Learning When Not to Decide: A Framework for Overcoming Factual Presumptuousness in AI Adjudication

Model ReleasesDGX agent

arXiv:2604.19895v1 Announce Type: new Abstract: A well-known limitation of AI systems is presumptuousness: the tendency of AI systems to provide confident answers when information may be lacking. This

Less Languages, Less Tokens: An Efficient Unified Logic Cross-lingual Chain-of-Thought Reasoning Framework

Model ReleasesDGX agent

arXiv:2604.20090v1 Announce Type: new Abstract: Cross-lingual chain-of-thought (XCoT) with self-consistency markedly enhances multilingual reasoning, yet existing methods remain costly due to extensiv

LEXIS: LatEnt ProXimal Interaction Signatures for 3D HOI from an Image

ResearchDGX agent

arXiv:2604.20800v1 Announce Type: new Abstract: Reconstructing 3D Human-Object Interaction from an RGB image is essential for perceptive systems. Yet, this remains challenging as it requires capturing

LLM Agents Grounded in Self-Reports Enable General-Purpose Simulation of Individuals

AgentsDGX agent

arXiv:2411.10109v2 Announce Type: replace Abstract: Machine learning can predict human behavior well when substantial structured data and well-defined outcomes are available, but these models are typi

LLM Agents Predict Social Media Reactions but Do Not Outperform Text Classifiers: Benchmarking Simulation Accuracy Using 120K+ Personas of 1511 Humans

SafetyDGX agent

arXiv:2604.19787v1 Announce Type: cross Abstract: Social media platforms mediate how billions form opinions and engage with public discourse. As autonomous AI agents increasingly participate in these

looks like new Pareto frontiers across everything: - Context: 400K context in Codex and a 1M in API - API Pricing: 5/m input and 30/m outp…

Model ReleasesDGX agent

looks like new Pareto frontiers across everything: - Context: 400K context in Codex and a 1M in API - API Pricing: 5/m input and 30/m output tokens. - Codex improved its own inference speed 20% lol -

MAPRPose: Mask-Aware Proposal and Amodal Refinement for Multi-Object 6D Pose Estimation

Model ReleasesDGX agent

arXiv:2604.20650v1 Announce Type: new Abstract: 6D object pose estimation in cluttered scenes remains challenging due to severe occlusion and sensor noise. We propose MAPRPose, a two-stage framework t

MetaboNet: The Largest Publicly Available Consolidated Dataset for Type 1 Diabetes Management

Model ReleasesDGX agent

arXiv:2601.11505v2 Announce Type: replace-cross Abstract: Progress in Type 1 Diabetes (T1D) algorithm development is limited by the fragmentation and lack of standardization across existing T1D manage

MGDA-Decoupled: Geometry-Aware Multi-Objective Optimisation for DPO-based LLM Alignment

SafetyDGX agent

arXiv:2604.20685v1 Announce Type: new Abstract: Aligning large language models (LLMs) to desirable human values requires balancing multiple, potentially conflicting objectives such as helpfulness, tru

🔹 One Prompt → 100-page PDF report +cited dataset + 30-page executive PPT + 20 financial charts

Model ReleasesDGX agent

Moonshot's Kimi AI demonstrated capabilities to generate comprehensive business reports from a single prompt, including a 100-page PDF with citations, a 30-page executive PowerPoint presentation, and

Optimal Single-Policy Sample Complexity and Transient Coverage for Average-Reward Offline RL

Model ReleasesDGX agent

arXiv:2506.20904v2 Announce Type: replace Abstract: We study offline reinforcement learning in average-reward MDPs, which presents increased challenges from the perspectives of distribution shift and

Optimizing User Profiles via Contextual Bandits for Retrieval-Augmented LLM Personalization

ResearchDGX agent

arXiv:2601.12078v2 Announce Type: replace Abstract: Large language models (LLMs) excel at general-purpose tasks, yet adapting their responses to individual users remains challenging. Retrieval augment

Over-Refusal and Representation Subspaces: A Mechanistic Analysis of Task-Conditioned Refusal in Aligned LLMs

ResearchDGX agent

arXiv:2603.27518v2 Announce Type: replace Abstract: Aligned language models that are trained to refuse harmful requests also exhibit over-refusal: they decline safe instructions that seemingly resembl

Over the past month, some of you reported Claude Code's quality had slipped. We investigated, and published a post-mortem on the three issue…

Model ReleasesDGX agent

Over the past month, some of you reported Claude Code's quality had slipped. We investigated, and published a post-mortem on the three issues we found. All are fixed in v2.1.116+ and we’ve reset usage

OVPD: A Virtual-Physical Fusion Testing Dataset of OnSite Auton-omous Driving Challenge

Model ReleasesDGX agent

arXiv:2604.20423v1 Announce Type: new Abstract: The rapid iteration of autonomous driving algorithms has created a growing demand for high-fidelity, replayable, and diagnosable testing data. However,

Pairing Regularization for Mitigating Many-to-One Collapse in GANs

ResearchDGX agent

arXiv:2604.20130v1 Announce Type: cross Abstract: Mode collapse remains a fundamental challenge in training generative adversarial networks (GANs). While existing works have primarily focused on inter

ParseBench is now live on @Kaggle. The first document OCR benchmark built for AI agents — 2,000 enterprise pages, 167K+ test rules, 5 dimens…

Model ReleasesDGX agent

ParseBench is now live on @Kaggle. The first document OCR benchmark built for AI agents — 2,000 enterprise pages, 167K+ test rules, 5 dimensions that actually break downstream agents. Benchmark your p

Personalized electric vehicle energy consumption estimation framework that integrates driver behavior with map data

ResearchDGX agent

arXiv:2604.20764v1 Announce Type: cross Abstract: This paper presents a personalized Battery Electric Vehicle (BEV) energy consumption estimation framework that integrates map-based contextual feature

Physics-Informed Conditional Diffusion for Motion-Robust Retinal Temporal Laser Speckle Contrast Imaging

ResearchDGX agent

arXiv:2604.20594v1 Announce Type: new Abstract: Retinal laser speckle contrast imaging (LSCI) is a noninvasive optical modality for monitoring retinal blood flow dynamics. However, conventional tempor

Portal26 launches Agentic Token Controls to cap runaway AI agent spend

Model ReleasesDGX agent

Generative artificial intelligence security startup Portal26 Inc. today announced the launch of a new module designed to rein in runaway token consumption by autonomous AI agents, a problem the compan

Prism: An Evolutionary Memory Substrate for Multi-Agent Open-Ended Discovery

Model ReleasesDGX agent

arXiv:2604.19795v1 Announce Type: new Abstract: We introduce prism{} (extbf{P}robabilistic extbf{R}etrieval with extbf{I}nformation-extbf{S}tratified extbf{M}emory), an evolutionary memory substrate f

Random Walk on Point Clouds for Feature Detection

ResearchDGX agent

arXiv:2604.20474v1 Announce Type: new Abstract: The points on the point clouds that can entirely outline the shape of the model are of critical importance, as they serve as the foundation for numerous

RefAerial: A Benchmark and Approach for Referring Detection in Aerial Images

Model ReleasesDGX agent

arXiv:2604.20543v1 Announce Type: new Abstract: Referring detection refers to locate the target referred by natural languages, which has recently attracted growing research interests. However, existin

Relative Entropy Estimation in Function Space: Theory and Applications to Trajectory Inference

Model ReleasesDGX agent

arXiv:2604.20775v1 Announce Type: new Abstract: Trajectory Inference (TI) seeks to recover latent dynamical processes from snapshot data, where only independent samples from time-indexed marginals are

Rethinking Reinforcement Fine-Tuning in LVLM: Convergence, Reward Decomposition, and Generalization

SafetyDGX agent

arXiv:2604.19857v1 Announce Type: cross Abstract: Reinforcement fine-tuning with verifiable rewards (RLVR) has emerged as a powerful paradigm for equipping large vision-language models (LVLMs) with ag

retinalysis-vascx: An explainable software toolbox for the extraction of retinal vascular biomarkers

Model ReleasesDGX agent

arXiv:2602.08580v2 Announce Type: replace-cross Abstract: Automatic extraction of retinal vascular biomarkers from color fundus images (CFI) is crucial for large-scale studies of the retinal vasculatu

Robustness of Spatio-temporal Graph Neural Networks for Fault Location in Partially Observable Distribution Grids

ResearchDGX agent

arXiv:2604.20403v1 Announce Type: new Abstract: Fault location in distribution grids is critical for reliability and minimizing outage durations. Yet, it remains challenging due to partial observabili

Rodrigues Network for Learning Robot Actions

SafetyDGX agent

arXiv:2506.02618v2 Announce Type: replace-cross Abstract: Understanding and predicting articulated actions is important in robot learning. However, common architectures such as MLPs and Transformers l

Same Content, Different Answers: Cross-Modal Inconsistency in MLLMs

ResearchDGX agent

arXiv:2512.08923v2 Announce Type: replace Abstract: We introduce two new benchmarks REST and REST+ (Render-Equivalence Stress Tests) to enable systematic evaluation of cross-modal inconsistency in mul

SceneOrchestra: Efficient Agentic 3D Scene Synthesis via Full Tool-Call Trajectory Generation

Model ReleasesDGX agent

arXiv:2604.19907v1 Announce Type: new Abstract: Recent agentic frameworks for 3D scene synthesis have advanced realism and diversity by integrating heterogeneous generation and editing tools. These to

Self-supervised pretraining for an iterative image size agnostic vision transformer

ResearchDGX agent

arXiv:2604.20392v1 Announce Type: new Abstract: Vision Transformers (ViTs) dominate self-supervised learning (SSL). While they have proven highly effective for large-scale pretraining, they are comput

SkillGraph: Graph Foundation Priors for LLM Agent Tool Sequence Recommendation

Model ReleasesDGX agent

arXiv:2604.19793v1 Announce Type: new Abstract: LLM agents must select tools from large API libraries and order them correctly. Existing methods use semantic similarity for both retrieval and ordering

Skyline-First Traversal as a Control Mechanism for Multi-Criteria Graph Search

ResearchDGX agent

arXiv:2604.19807v1 Announce Type: new Abstract: In multi-criteria graph traversal, paths are compared via Pareto dominance, an ordering that identifies which paths are non-dominated, but says nothing

Snow Flurries: How UNC6692 Employed Social Engineering to Deploy a Custom Malware Suite

Model ReleasesDGX agent

Written by: JP Glab, Tufail Ahmed, Josh Kelley, Muhammad Umair Introduction Google Threat Intelligence Group (GTIG) identified a multistage intrusion campaign by a newly tracked threat group, UNC6692,

Soft-Label Governance for Distributional Safety in Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2604.19752v1 Announce Type: cross Abstract: Multi-agent AI systems exhibit emergent risks that no single agent produces in isolation. Existing safety frameworks rely on binary classifications of

← Previous
1…708709710711712…1035
Next →