AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

Content type
AllBlog
87,042Total entries
1Added by human
87,041Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,509 results
Model Releases

Fast Bayesian equipment condition monitoring via simulation based inference: applications to heat exchanger health

DGX agent

arXiv:2604.20735v1 Announce Type: new Abstract: Accurate condition monitoring of industrial equipment requires inferring latent degradation parameters from indirect sensor measurements under uncertain

model-releasesarxiv-cs-lg
23 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Finding Duplicates in 1.1M BDD Steps: cukereuse, a Paraphrase-Robust Static Detector for Cucumber and Gherkin

DGX agent

arXiv:2604.20462v1 Announce Type: cross Abstract: Behaviour-Driven Development (BDD) suites accumulate step-text duplication whose maintenance cost is established in prior work. Existing detection tec

model-releasesarxiv-cs-cl
23 Apr 2026
Agents

Forage V2: Knowledge Evolution and Transfer in Autonomous Agent Organizations

DGX agent

arXiv:2604.19837v1 Announce Type: new Abstract: Autonomous agents operating in open-world tasks -- where the completion boundary is not given in advance -- face denominator blindness: they systematica

agentsarxiv-cs-ai
23 Apr 2026
Model Releases

From Recall to Forgetting: Benchmarking Long-Term Memory for Personalized Agents

DGX agent

arXiv:2604.20006v1 Announce Type: new Abstract: Personalized agents that interact with users over long periods must maintain persistent memory across sessions and update it as circumstances change. Ho

model-releasesarxiv-cs-cl
23 Apr 2026
Research

Gauge-covariant stochastic neural fields: Stability and finite-width effects

DGX agent

arXiv:2508.18948v2 Announce Type: replace-cross Abstract: We develop a gauge-covariant stochastic effective field theory for stability and finite-width effects in deep neural systems. The model uses c

researcharxiv-cs-lg
23 Apr 2026
Model Releases

Global Offshore Wind Infrastructure: Deployment and Operational Dynamics from Dense Sentinel-1 Time Series

DGX agent

arXiv:2604.20822v1 Announce Type: new Abstract: The offshore wind energy sector is expanding rapidly, increasing the need for independent, high-temporal-resolution monitoring of infrastructure deploym

model-releasesarxiv-cs-cv
23 Apr 2026
Model Releases

GPT-5.5 Bio Bug Bounty

DGX agent

OpenAI's GPT-5.5 Bio Bug Bounty program invites security researchers to identify and report vulnerabilities in GPT-5.5's biological information handling capabilities, focusing on potential misuse risk

model-releasesopenai
23 Apr 2026
Model Releases

GPT-5.5 is now accessible in Hermes Agent through the ChatGPT/Codex OAuth provider. Run `hermes update` to access now or learn how to get st…

DGX agent

GPT-5.5 is now accessible in Hermes Agent through the ChatGPT/Codex OAuth provider. Run `hermes update` to access now or learn how to get started with Hermes Agent here: https://hermes-agent.nousresea

model-releasesnous-research--x
23 Apr 2026
Model Releases

GPT-5.5 may not be in the official OpenAI API... but it's available via the apparently approved-of Codex API backdoor So I used that to make…

DGX agent

GPT-5.5 may not be in the official OpenAI API... but it's available via the apparently approved-of Codex API backdoor So I used that to make these pelicans (default and xhigh)! https://simonwillison.n

model-releasessimon-willison--x
23 Apr 2026
Safety

Hybrid Latent Reasoning with Decoupled Policy Optimization

DGX agent

arXiv:2604.20328v1 Announce Type: new Abstract: Chain-of-Thought (CoT) reasoning significantly elevates the complex problem-solving capabilities of multimodal large language models (MLLMs). However, a

safetyarxiv-cs-cv
23 Apr 2026
Model Releases

Hybrid Multi-Phase Page Matching and Multi-Layer Diff Detection for Japanese Building Permit Document Review

DGX agent

arXiv:2604.19770v1 Announce Type: new Abstract: We present a hybrid multi-phase page matching algorithm for automated comparison of Japanese building permit document sets. Building permit review in Ja

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

I had early access to GPT-5.5. It is very good, especially the Pro version. Full writeup very shortly.

DGX agent

Ethan Mollick posted on X about having early access to GPT-5.5, commenting positively on its capabilities and noting that the Pro version is particularly strong. He indicated that a full detailed writ

model-releasesethan-mollick--x
23 Apr 2026
Model Releases

I’d been part of OpenAI early tester group for GPT-5.5. I believe with GPT-5.5 Pro we reached another inflection point-comparable to the ori…

DGX agent

I’d been part of OpenAI early tester group for GPT-5.5. I believe with GPT-5.5 Pro we reached another inflection point-comparable to the original release of o1-preview & then with 5.0 Pro, I had felt.

model-releasessam-altman--x
23 Apr 2026
Model Releases

If you want to stack rank LLMs/VLMs on document understanding 📄, you can through ParseBench, now live on @kaggle 📊 ParseBench is the most …

DGX agent

If you want to stack rank LLMs/VLMs on document understanding 📄, you can through ParseBench, now live on @kaggle 📊 ParseBench is the most comprehensive document OCR benchmark over real enterprise docu

model-releasesjerry-liu--x
23 Apr 2026
Agents

If you're waiting for a sign... that might not be it! Mitigating Trust Boundary Confusion from Visual Injections on Vision-Language Agentic Systems

DGX agent

arXiv:2604.19844v1 Announce Type: cross Abstract: Recent advances in embodied Vision-Language Agentic Systems (VLAS), powered by large vision-language models (LVLMs), enable AI systems to perceive and

agentsarxiv-cs-ai
23 Apr 2026
Model Releases

IMPACT-CYCLE: A Contract-Based Multi-Agent System for Claim-Level Supervisory Correction of Long-Video Semantic Memory

DGX agent

arXiv:2604.20136v1 Announce Type: cross Abstract: Correcting errors in long-video understanding is disproportionately costly: existing multimodal pipelines produce opaque, end-to-end outputs that expo

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

🚨 In a new court filing (below), Clippers owner Steve Ballmer dismisses @pablofindsout as “gossip” from a “former talking head and televisi…

DGX agent

🚨 In a new court filing (below), Clippers owner Steve Ballmer dismisses @pablofindsout as “gossip” from a “former talking head and television personality.” Here is an excerpt from the federal whistleb

model-releasesanthropic--x
23 Apr 2026
Agents

In the enterprise AI race, who is leading and who is just reacting?

DGX agent

Enterprise AI scaling is accelerating as organizations shift from experimentation to full deployment, embedding intelligence into core workflows. At the same time, agentic systems are driving a broade

agentssiliconangle
23 Apr 2026
Model Releases

Instagram launches Instants, an app for sharing disappearing photos, in Italy and Spain, after rolling out an Instants feature in its main app in some regions (Sydney Bradley/Business Insider)

DGX agent

Sydney Bradley / Business Insider: Instagram launches Instants, an app for sharing disappearing photos, in Italy and Spain, after rolling out an Instants feature in its main app in some regions — - In

model-releasestechmeme
23 Apr 2026
Model Releases

Interesting, OpenAI just released a free healthcare version of ChatGPT-5.4 for clinicians that beat specialty-matched physicians with unlimi…

DGX agent

Interesting, OpenAI just released a free healthcare version of ChatGPT-5.4 for clinicians that beat specialty-matched physicians with unlimited time + web access on a benchmark of real & hard clinical

model-releasesethan-mollick--x
23 Apr 2026
Model Releases

Introducing GPT-5.5 A new class of intelligence for real work and powering agents, built to understand complex goals, use tools, check its w…

DGX agent

Introducing GPT-5.5 A new class of intelligence for real work and powering agents, built to understand complex goals, use tools, check its work, and carry more tasks through to completion. It marks a

model-releasesopenai--x
23 Apr 2026
Local Ai

Klein 9b base nvfp4 on HF

DGX agent

FLUX.2 [klein] 9B Base is a 9 billion parameter rectified flow transformer capable of generating images from text descriptions and supports multi-reference editing capabilities. The model fits in appr

local-air-stablediffusion
23 Apr 2026
Local Ai

Learning Spatial-Temporal Coherent Correlations for Speech-Preserving Facial Expression Manipulation

DGX agent

arXiv:2604.20226v1 Announce Type: new Abstract: Speech-preserving facial expression manipulation (SPFEM) aims to modify facial emotions while meticulously maintaining the mouth animation associated wi

local-aiarxiv-cs-cv
23 Apr 2026
Model Releases

Learning When Not to Decide: A Framework for Overcoming Factual Presumptuousness in AI Adjudication

DGX agent

arXiv:2604.19895v1 Announce Type: new Abstract: A well-known limitation of AI systems is presumptuousness: the tendency of AI systems to provide confident answers when information may be lacking. This

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Less Languages, Less Tokens: An Efficient Unified Logic Cross-lingual Chain-of-Thought Reasoning Framework

DGX agent

arXiv:2604.20090v1 Announce Type: new Abstract: Cross-lingual chain-of-thought (XCoT) with self-consistency markedly enhances multilingual reasoning, yet existing methods remain costly due to extensiv

model-releasesarxiv-cs-cl
23 Apr 2026
Research

LEXIS: LatEnt ProXimal Interaction Signatures for 3D HOI from an Image

DGX agent

arXiv:2604.20800v1 Announce Type: new Abstract: Reconstructing 3D Human-Object Interaction from an RGB image is essential for perceptive systems. Yet, this remains challenging as it requires capturing

researcharxiv-cs-cv
23 Apr 2026
Agents

LLM Agents Grounded in Self-Reports Enable General-Purpose Simulation of Individuals

DGX agent

arXiv:2411.10109v2 Announce Type: replace Abstract: Machine learning can predict human behavior well when substantial structured data and well-defined outcomes are available, but these models are typi

agentsarxiv-cs-ai
23 Apr 2026
Safety

LLM Agents Predict Social Media Reactions but Do Not Outperform Text Classifiers: Benchmarking Simulation Accuracy Using 120K+ Personas of 1511 Humans

DGX agent

arXiv:2604.19787v1 Announce Type: cross Abstract: Social media platforms mediate how billions form opinions and engage with public discourse. As autonomous AI agents increasingly participate in these

safetyarxiv-cs-ai
23 Apr 2026
Model Releases

looks like new Pareto frontiers across everything: - Context: 400K context in Codex and a 1M in API - API Pricing: 5/m input and 30/m outp…

DGX agent

looks like new Pareto frontiers across everything: - Context: 400K context in Codex and a 1M in API - API Pricing: 5/m input and 30/m output tokens. - Codex improved its own inference speed 20% lol -

model-releasesswyx--x
23 Apr 2026
Model Releases

MAPRPose: Mask-Aware Proposal and Amodal Refinement for Multi-Object 6D Pose Estimation

DGX agent

arXiv:2604.20650v1 Announce Type: new Abstract: 6D object pose estimation in cluttered scenes remains challenging due to severe occlusion and sensor noise. We propose MAPRPose, a two-stage framework t

model-releasesarxiv-cs-cv
23 Apr 2026
Model Releases

MetaboNet: The Largest Publicly Available Consolidated Dataset for Type 1 Diabetes Management

DGX agent

arXiv:2601.11505v2 Announce Type: replace-cross Abstract: Progress in Type 1 Diabetes (T1D) algorithm development is limited by the fragmentation and lack of standardization across existing T1D manage

model-releasesarxiv-cs-ai
23 Apr 2026
Safety

MGDA-Decoupled: Geometry-Aware Multi-Objective Optimisation for DPO-based LLM Alignment

DGX agent

arXiv:2604.20685v1 Announce Type: new Abstract: Aligning large language models (LLMs) to desirable human values requires balancing multiple, potentially conflicting objectives such as helpfulness, tru

safetyarxiv-cs-lg
23 Apr 2026
Model Releases

🔹 One Prompt → 100-page PDF report +cited dataset + 30-page executive PPT + 20 financial charts

DGX agent

Moonshot's Kimi AI demonstrated capabilities to generate comprehensive business reports from a single prompt, including a 100-page PDF with citations, a 30-page executive PowerPoint presentation, and

model-releaseskimi-moonshot--x
23 Apr 2026
Model Releases

Optimal Single-Policy Sample Complexity and Transient Coverage for Average-Reward Offline RL

DGX agent

arXiv:2506.20904v2 Announce Type: replace Abstract: We study offline reinforcement learning in average-reward MDPs, which presents increased challenges from the perspectives of distribution shift and

model-releasesarxiv-cs-lg
23 Apr 2026
Research

Optimizing User Profiles via Contextual Bandits for Retrieval-Augmented LLM Personalization

DGX agent

arXiv:2601.12078v2 Announce Type: replace Abstract: Large language models (LLMs) excel at general-purpose tasks, yet adapting their responses to individual users remains challenging. Retrieval augment

researcharxiv-cs-cl
23 Apr 2026
Research

Over-Refusal and Representation Subspaces: A Mechanistic Analysis of Task-Conditioned Refusal in Aligned LLMs

DGX agent

arXiv:2603.27518v2 Announce Type: replace Abstract: Aligned language models that are trained to refuse harmful requests also exhibit over-refusal: they decline safe instructions that seemingly resembl

researcharxiv-cs-cl
23 Apr 2026
Model Releases

Over the past month, some of you reported Claude Code's quality had slipped. We investigated, and published a post-mortem on the three issue…

DGX agent

Over the past month, some of you reported Claude Code's quality had slipped. We investigated, and published a post-mortem on the three issues we found. All are fixed in v2.1.116+ and we’ve reset usage

model-releasesthariq--x
23 Apr 2026
Model Releases

OVPD: A Virtual-Physical Fusion Testing Dataset of OnSite Auton-omous Driving Challenge

DGX agent

arXiv:2604.20423v1 Announce Type: new Abstract: The rapid iteration of autonomous driving algorithms has created a growing demand for high-fidelity, replayable, and diagnosable testing data. However,

model-releasesarxiv-cs-ro
23 Apr 2026
Research

Pairing Regularization for Mitigating Many-to-One Collapse in GANs

DGX agent

arXiv:2604.20130v1 Announce Type: cross Abstract: Mode collapse remains a fundamental challenge in training generative adversarial networks (GANs). While existing works have primarily focused on inter

researcharxiv-cs-cv
23 Apr 2026
Model Releases

ParseBench is now live on @Kaggle. The first document OCR benchmark built for AI agents — 2,000 enterprise pages, 167K+ test rules, 5 dimens…

DGX agent

ParseBench is now live on @Kaggle. The first document OCR benchmark built for AI agents — 2,000 enterprise pages, 167K+ test rules, 5 dimensions that actually break downstream agents. Benchmark your p

model-releasesjerry-liu--x
23 Apr 2026
Research

Personalized electric vehicle energy consumption estimation framework that integrates driver behavior with map data

DGX agent

arXiv:2604.20764v1 Announce Type: cross Abstract: This paper presents a personalized Battery Electric Vehicle (BEV) energy consumption estimation framework that integrates map-based contextual feature

researcharxiv-cs-lg
23 Apr 2026
Research

Physics-Informed Conditional Diffusion for Motion-Robust Retinal Temporal Laser Speckle Contrast Imaging

DGX agent

arXiv:2604.20594v1 Announce Type: new Abstract: Retinal laser speckle contrast imaging (LSCI) is a noninvasive optical modality for monitoring retinal blood flow dynamics. However, conventional tempor

researcharxiv-cs-cv
23 Apr 2026
Model Releases

Portal26 launches Agentic Token Controls to cap runaway AI agent spend

DGX agent

Generative artificial intelligence security startup Portal26 Inc. today announced the launch of a new module designed to rein in runaway token consumption by autonomous AI agents, a problem the compan

model-releasessiliconangle
23 Apr 2026
Model Releases

Prism: An Evolutionary Memory Substrate for Multi-Agent Open-Ended Discovery

DGX agent

arXiv:2604.19795v1 Announce Type: new Abstract: We introduce prism{} (extbf{P}robabilistic extbf{R}etrieval with extbf{I}nformation-extbf{S}tratified extbf{M}emory), an evolutionary memory substrate f

model-releasesarxiv-cs-ai
23 Apr 2026
Research

Random Walk on Point Clouds for Feature Detection

DGX agent

arXiv:2604.20474v1 Announce Type: new Abstract: The points on the point clouds that can entirely outline the shape of the model are of critical importance, as they serve as the foundation for numerous

researcharxiv-cs-cv
23 Apr 2026
Model Releases

RefAerial: A Benchmark and Approach for Referring Detection in Aerial Images

DGX agent

arXiv:2604.20543v1 Announce Type: new Abstract: Referring detection refers to locate the target referred by natural languages, which has recently attracted growing research interests. However, existin

model-releasesarxiv-cs-cv
23 Apr 2026
Model Releases

Relative Entropy Estimation in Function Space: Theory and Applications to Trajectory Inference

DGX agent

arXiv:2604.20775v1 Announce Type: new Abstract: Trajectory Inference (TI) seeks to recover latent dynamical processes from snapshot data, where only independent samples from time-indexed marginals are

model-releasesarxiv-cs-lg
23 Apr 2026
Safety

Rethinking Reinforcement Fine-Tuning in LVLM: Convergence, Reward Decomposition, and Generalization

DGX agent

arXiv:2604.19857v1 Announce Type: cross Abstract: Reinforcement fine-tuning with verifiable rewards (RLVR) has emerged as a powerful paradigm for equipping large vision-language models (LVLMs) with ag

safetyarxiv-cs-cl
23 Apr 2026
← Previous
1…892893894895896…1303
Next →