AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
9,955 results
24 Apr 2026

Q&A with Google Cloud CEO Thomas Kurian on building infrastructure for AI agents, balancing internal needs and the demands of customers like Anthropic, and more (Ben Thompson/Stratechery)

AgentsDGX agent

Ben Thompson / Stratechery: Q&A with Google Cloud CEO Thomas Kurian on building infrastructure for AI agents, balancing internal needs and the demands of customers like Anthropic, and more — This week

Read More: https://blog.comfy.org/p/comfyui-raises-30m-to-scale-open

Local AiDGX agent

ComfyUI announced a $30 million funding round to support scaling and development of its open-source node-based UI platform for AI image generation. The funding will likely be used to expand infrastruc

RELOOP: Recursive Retrieval with Multi-Hop Reasoner and Planners for Heterogeneous QA

SafetyDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2510.20505v4 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) remains brittle on multi-step questions and heterogeneous evidence sources, trading accuracy against late

Serialisation Strategy Matters: How FHIR Data Format Affects LLM Medication Reconciliation

Model ReleasesDGX agent

arXiv:2604.21076v1 Announce Type: cross Abstract: Medication reconciliation at clinical handoffs is a high-stakes, error-prone process. Large language models are increasingly proposed to assist with t

Structural Quality Gaps in Practitioner AI Governance Prompts: An Empirical Study Using a Five-Principle Evaluation Framework

AgentsDGX agent

arXiv:2604.21090v1 Announce Type: cross Abstract: AI governance programmes increasingly rely on natural language prompts to constrain and direct AI agent behaviour. These prompts function as executabl

SurgViVQA: Temporally-Grounded Video Question Answering for Surgical Scene Understanding

Model ReleasesDGX agent

arXiv:2511.03325v3 Announce Type: replace Abstract: Video Question Answering (VideoQA) in the surgical domain aims to enhance intraoperative understanding by enabling AI models to reason over temporal

this was a good week. proud of the team. happy building!

IndustryDGX agent

Sam Altman shared a brief positive message on X expressing satisfaction with recent team accomplishments and progress. The post reflects optimism about ongoing projects and company morale at OpenAI.

Tracxn: global edtech funding fell from 16.7B in 2021 to 2.6B in 2025, while the number of startups launched dropped from 10,491 in 2021 to just 645 in 2025 (Ananya Bhattacharya/Rest of World)

IndustryDGX agent

Ananya Bhattacharya / Rest of World: Tracxn: global edtech funding fell from 16.7B in 2021 to 2.6B in 2025, while the number of startups launched dropped from 10,491 in 2021 to just 645 in 2025 — Vent

Value-Conflict Diagnostics Reveal Widespread Alignment Faking in Language Models

SafetyDGX agent

arXiv:2604.20995v1 Announce Type: new Abstract: Alignment faking, where a model behaves aligned with developer policy when monitored but reverts to its own preferences when unobserved, is a concerning

Very cool demo by @Box how progressive skill loading works in practice 👀 The agent walks through onboarding step-by-step: load employee → l…

AgentsDGX agent

Very cool demo by @Box how progressive skill loading works in practice 👀 The agent walks through onboarding step-by-step: load employee → load skill → scoped search → synthesize Instead of one giant p

VLAA-GUI: Knowing When to Stop, Recover, and Search, A Modular Framework for GUI Automation

Model ReleasesDGX agent

arXiv:2604.21375v1 Announce Type: cross Abstract: Autonomous GUI agents face two fundamental challenges: early stopping, where agents prematurely declare success without verifiable evidence, and repet

We’ve refreshed Claude Code on the web and mobile. A few things that recently shipped 🧵

Model ReleasesDGX agent

Claude Code received recent updates on both web and mobile platforms, with new features and improvements announced in a thread format. The refresh likely includes interface enhancements, performance i

Wiring the 'Why': A Unified Taxonomy and Survey of Abductive Reasoning in LLMs

Model ReleasesDGX agent

arXiv:2604.08016v2 Announce Type: replace Abstract: Regardless of its foundational role in human discovery and sense-making, abductive reasoning--the inference of the most plausible explanation for an

23 Apr 2026

A Computational Model of Message Sensation Value in Short Video Multimodal Features that Predicts Sensory and Behavioral Engagement

ResearchDGX agent

arXiv:2604.19995v1 Announce Type: new Abstract: The contemporary media landscape is characterized by sensational short videos. While prior research examines the effects of individual multimodal featur

A Field Guide to Decision Making

AgentsDGX agent

arXiv:2604.20669v1 Announce Type: cross Abstract: High-consequence decision making demands peak performance from individuals in positions of responsibility. Such executive authority bears the obligati

A Kinematic Framework for Evaluating Pinch Configurations in Robotic Hand Design without Object or Contact Models

ResearchDGX agent

arXiv:2604.20692v1 Announce Type: new Abstract: Evaluating the pinch capability of a robotic hand is important for understanding its functional dexterity. However, many existing grasp evaluation metho

Accumulated Aggregated D-Optimal Designs for Estimating Main Effects in Black-Box Models

ResearchDGX agent

arXiv:2510.08465v2 Announce Type: replace-cross Abstract: Estimating how individual input variables affect the output of a black-box model is a central task in explainable machine learning. However, e

Auto-Unrolled Proximal Gradient Descent: An AutoML Approach to Interpretable Waveform Optimization

ResearchDGX agent

arXiv:2603.17478v2 Announce Type: replace-cross Abstract: This study explores the combination of automated machine learning (AutoML) with model-based deep unfolding (DU) for optimizing wireless beamfo

Automatic Ontology Construction Using LLMs as an External Layer of Memory, Verification, and Planning for Hybrid Intelligent Systems

Model ReleasesDGX agent

arXiv:2604.20795v1 Announce Type: new Abstract: This paper presents a hybrid architecture for intelligent systems in which large language models (LLMs) are extended with an external ontological memory

@benswerd there are tradeoffs but we've spoken to folks doing both inside and outside and the split is pretty much 50:50 in our experience. …

AgentsDGX agent

@benswerd there are tradeoffs but we've spoken to folks doing both inside and outside and the split is pretty much 50:50 in our experience. I've found @hwchase17's list of tradeoffs the most accurate

Beyond Text-Dominance: Understanding Modality Preference of Omni-modal Large Language Models

Model ReleasesDGX agent

arXiv:2604.16902v2 Announce Type: replace Abstract: Native Omni-modal Large Language Models (OLLMs) have shifted from pipeline architectures to unified representation spaces. However, this native inte

btw in talking to friends the best framing for how to discuss GPT-Image-2-Thinking taking multiple tens of mins for generation and being abl…

Model ReleasesDGX agent

btw in talking to friends the best framing for how to discuss GPT-Image-2-Thinking taking multiple tens of mins for generation and being able to oneshot QR codes and diagrams and logos and foods and f

Causal-Transformer with Adaptive Mutation-Locking for Early Prediction of Acute Kidney Injury

ResearchDGX agent

arXiv:2604.20259v1 Announce Type: new Abstract: Accurate early prediction of Acute Kidney Injury (AKI) is critical for timely clinical intervention. However, existing deep learning models struggle wit

Cognitive Kernel-Pro: A Framework for Deep Research Agents and Agent Foundation Models Training

Model ReleasesDGX agent

arXiv:2508.00414v3 Announce Type: replace Abstract: General AI Agents are increasingly recognized as foundational frameworks for the next generation of artificial intelligence, enabling complex reason

Copperhelm dives deep into automation to build enterprise cloud defenses

AgentsDGX agent

Copperhelm Inc., a startup building agentic artificial intelligence for cloud cybersecurity, today announced its launch with $7 million in seed funding led by TLV Partners. ToDay Ventures, ICON and Sa

CreativeGame:Toward Mechanic-Aware Creative Game Generation

AgentsDGX agent

arXiv:2604.19926v1 Announce Type: new Abstract: Large language models can generate plausible game code, but turning this capability into iterative creative improvement remains difficult. In practice,

Don't burn tokens on search. 📉🔥 Will Templeton, CTO and co-founder at Allspice, explains why they don't let LLMs search their 10,000+ ingr…

ApplicationsDGX agent

Don't burn tokens on search. 📉🔥 Will Templeton, CTO and co-founder at Allspice, explains why they don't let LLMs search their 10,000+ ingredient database: ❌ Searching 10k items in a prompt = waste of

Evals ~= Environments…they’re one of the best investments a team can make for improving agents Step 0: Turn On Tracing for Agents Step 1: Po…

AgentsDGX agent

Evals ~= Environments…they’re one of the best investments a team can make for improving agents Step 0: Turn On Tracing for Agents Step 1: Point compute at Traces to understand agent behavior, segment

Explicit Dropout: Deterministic Regularization for Transformer Architectures

ResearchDGX agent

arXiv:2604.20505v1 Announce Type: new Abstract: Dropout is a widely used regularization technique in deep learning, but its effects are typically realized through stochastic masking rather than explic

Finding Duplicates in 1.1M BDD Steps: cukereuse, a Paraphrase-Robust Static Detector for Cucumber and Gherkin

Model ReleasesDGX agent

arXiv:2604.20462v1 Announce Type: cross Abstract: Behaviour-Driven Development (BDD) suites accumulate step-text duplication whose maintenance cost is established in prior work. Existing detection tec

From Data to Theory: Autonomous Large Language Model Agents for Materials Science

Model ReleasesDGX agent

arXiv:2604.19789v1 Announce Type: new Abstract: We present an autonomous large language model (LLM) agent for end-to-end, data-driven materials theory development. The model can choose an equation for

From Fuzzy to Formal: Scaling Hospital Quality Improvement with AI

SafetyDGX agent

arXiv:2604.20055v1 Announce Type: new Abstract: Hospital Quality Improvement (QI) plays a critical role in optimizing healthcare delivery by translating high-level hospital goals into actionable solut

From GPUs to AI factories: Inside the Nvidia-Google Cloud superstack

HardwareDGX agent

Nvidia Corp. and Google LLC used the search giant’s annual Cloud Next event to deepen their long-running partnership, creating a full-stack “artificial intelligence factory” that integrates Google’s A

GPT-5.5 in Codex is a delight to work with: - Super sharp with responses - It understands intent better than any model - Great 'personality'…

Model ReleasesDGX agent

GPT-5.5 in Codex is a delight to work with: - Super sharp with responses - It understands intent better than any model - Great 'personality' - Gets lots of stuff done without pausing unnecessarily It

GPT-5.5 is now accessible in Hermes Agent through the ChatGPT/Codex OAuth provider. Run `hermes update` to access now or learn how to get st…

Model ReleasesDGX agent

GPT-5.5 is now accessible in Hermes Agent through the ChatGPT/Codex OAuth provider. Run `hermes update` to access now or learn how to get started with Hermes Agent here: https://hermes-agent.nousresea

GPT-5.5 is rolling out to Plus, Pro, Business, and Enterprise users in ChatGPT and Codex, and GPT-5.5 Pro to Pro, Business, and Enterprise users in ChatGPT (The Verge)

Model ReleasesDGX agent

The Verge: GPT-5.5 is rolling out to Plus, Pro, Business, and Enterprise users in ChatGPT and Codex, and GPT-5.5 Pro to Pro, Business, and Enterprise users in ChatGPT — The new model ‘excels’ at tasks

Graph-Theoretic Models for the Prediction of Molecular Measurements

Model ReleasesDGX agent

arXiv:2604.19840v1 Announce Type: new Abstract: Graph-theoretic approaches offer simplicity, interpretability, and low computational cost for molecular property prediction. Among these, the model prop

https://cognition.ai/blog/what-we-learned-building-cloud-agents

AgentsDGX agent

Cognition AI shares insights and lessons learned from developing cloud-based AI agents, likely covering technical challenges, architectural decisions, and practical implementation strategies for build

Human-like Content Analysis for Generative AI with Language-Grounded Sparse Encoders

ApplicationsDGX agent

arXiv:2508.18236v4 Announce Type: replace Abstract: The rapid development of generative AI has transformed content creation, communication, and human development. However, this technology raises profo

Hybrid Latent Reasoning with Decoupled Policy Optimization

SafetyDGX agent

arXiv:2604.20328v1 Announce Type: new Abstract: Chain-of-Thought (CoT) reasoning significantly elevates the complex problem-solving capabilities of multimodal large language models (MLLMs). However, a

Is Automatic1111 still valid?

Local AiDGX agent

AUTOMATIC1111 remains a popular Stable Diffusion interface, though many advanced users have moved to alternatives like ComfyUI, Forge, or ReForge for better performance with newer models like FLUX and

I've been previewing this in Codex for a few weeks - it's very good! Had some great results from it having it run security reviews against c…

Model ReleasesDGX agent

I've been previewing this in Codex for a few weeks - it's very good! Had some great results from it having it run security reviews against code written using other models Introducing GPT-5.5 A new cla

Klein 9b base nvfp4 on HF

Local AiDGX agent

FLUX.2 [klein] 9B Base is a 9 billion parameter rectified flow transformer capable of generating images from text descriptions and supports multi-reference editing capabilities. The model fits in appr

Last week, we launched Gemini 3.1 TTS, our latest and best text-to-speech model. This new model introduces [awe] audio tags, an intuitive wa…

Model ReleasesDGX agent

Last week, we launched Gemini 3.1 TTS, our latest and best text-to-speech model. This new model introduces [awe] audio tags, an intuitive way to guide vocal style, pace, and delivery. Here are some ti

Learning to Evolve: A Self-Improving Framework for Multi-Agent Systems via Textual Parameter Graph Optimization

Model ReleasesDGX agent

arXiv:2604.20714v1 Announce Type: new Abstract: Designing and optimizing multi-agent systems (MAS) is a complex, labor-intensive process of 'Agent Engineering.' Existing automatic optimization methods

looks like new Pareto frontiers across everything: - Context: 400K context in Codex and a 1M in API - API Pricing: 5/m input and 30/m outp…

Model ReleasesDGX agent

looks like new Pareto frontiers across everything: - Context: 400K context in Codex and a 1M in API - API Pricing: 5/m input and 30/m output tokens. - Codex improved its own inference speed 20% lol -

Markov reads Pushkin, again: A statistical journey into the poetic world of Evgenij Onegin

ResearchDGX agent

arXiv:2604.20221v1 Announce Type: new Abstract: This study applies symbolic time series analysis and Markov modeling to explore the phonological structure of Evgenij Onegin-as captured through a graph

Model Capability Assessment and Safeguards for Biological Weaponization

Model ReleasesDGX agent

arXiv:2604.19811v1 Announce Type: cross Abstract: AI leaders and safety reports increasingly warn that advances in model reasoning may enable biological misuse, including by low-expertise users, while

Multi-Objective Reinforcement Learning for Generating Covalent Inhibitor Candidates

SafetyDGX agent

arXiv:2604.20019v1 Announce Type: new Abstract: Rational design of covalent inhibitors requires simultaneously optimizing multiple properties, such as binding affinity, target selectivity, or electrop

New in the Codex app: - GPT-5.5 - Browser control - Sheets & Slides - Docs & PDFs - OS-wide dictation - Auto-review mode Enjoy!

Model ReleasesDGX agent

The Codex app now includes several new features: GPT-5.5 integration, browser control capabilities, support for Google Sheets and Slides, document and PDF handling, OS-wide dictation functionality, an

New subscription tiers are live on Nous Portal → Plus (20) → Super (100) → Ultra ($200) Bonus credits on signups, upgrades, and renewals: …

ResearchDGX agent

New subscription tiers are live on Nous Portal → Plus (20) → Super (100) → Ultra (200) Bonus credits on signups, upgrades, and renewals: +2 on Plus / +10 on Super / +20 on Ultra All tiers include acce

Omni, which is building a 'semantic translation layer' for enterprise data, raised a 120M Series C led by Iconiq at a 1.5B valuation, up from $650M in 2025 (Lily Mae Lazarus/Fortune)

ApplicationsDGX agent

Lily Mae Lazarus / Fortune: Omni, which is building a “semantic translation layer” for enterprise data, raised a 120M Series C led by Iconiq at a 1.5B valuation, up from $650M in 2025 — For years, com

Open-source agent for long-horizon deep research https://github.com/TIGER-AI-Lab/OpenResearcher

AgentsDGX agent

OpenResearcher is an open-source AI agent designed to conduct long-horizon deep research tasks, enabling autonomous investigation and analysis across extended research workflows. The project is mainta

🚨 OpenAI just launched GPT-5.5. The OpenAI team was nice enough to give me early access over the last several weeks, and I just want to fla…

Model ReleasesDGX agent

🚨 OpenAI just launched GPT-5.5. The OpenAI team was nice enough to give me early access over the last several weeks, and I just want to flag: there is a certain class of models (one that we’re hitting

Participatory provenance as representational auditing for AI-mediated public consultation

SafetyDGX agent

arXiv:2604.20711v1 Announce Type: new Abstract: Artificial intelligence is increasingly deployed to synthesize large-scale public input in policy consultations and participatory processes. Yet no form

Plugins and skills

TutorialsDGX agent

Plugins and skills are extensions that enhance the capabilities of OpenAI's Codex, allowing developers to integrate custom functionality and extend the model's functionality beyond its base capabiliti

Rabies diagnosis in low-data settings: A comparative study on the impact of data augmentation and transfer learning

ResearchDGX agent

arXiv:2604.19823v1 Announce Type: cross Abstract: Rabies remains a major public health concern across many African and Asian countries, where accurate diagnosis is critical for effective epidemiologic

Really excellent work by the inference team to serve this model so efficiently! To a significant degree, we have to become an AI inference c…

IndustryDGX agent

Sam Altman praises the inference team's work on efficiently serving a model, suggesting that becoming proficient in AI inference is crucial to the field's progress. The post appears to highlight the t

Saying More Than They Know: A Framework for Quantifying Epistemic-Rhetorical Miscalibration in Large Language Models

ResearchDGX agent

arXiv:2604.19768v1 Announce Type: cross Abstract: Large language models (LLMs) exhibit systematic miscalibration with rhetorical intensity not proportionate to epistemic grounding. This study tests th

SWE-chat: Coding Agent Interactions From Real Users in the Wild

AgentsDGX agent

arXiv:2604.20779v1 Announce Type: new Abstract: AI coding agents are being adopted at scale, yet we lack empirical evidence on how people actually use them and how much of their output is useful in pr

← Previous
1…152153154155156…166
Next →