AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,599 results
24 Apr 2026

ComfyUI, which gives creators granular control over image, video, and audio outputs from diffusion models, raised 30M at a 500M valuation (Marina Temkin/TechCrunch)

IndustryDGX agent

Marina Temkin / TechCrunch: ComfyUI, which gives creators granular control over image, video, and audio outputs from diffusion models, raised 30M at a 500M valuation — ComfyUI, a startup that helps cr

Congrats to @deepseek_ai team! Doing the numbers I would estimate: Pro < 14m for the final training run Flash < 4m Ratio of active params …

Model ReleasesDGX agent

Congrats to @deepseek_ai team! Doing the numbers I would estimate: Pro < 14m for the final training run Flash < 4m Ratio of active params x total training tokens vs v3 Total compute costs (data prep,

Diplomatic cable: US State Department has ordered a global push to bring attention to what it says are efforts by Chinese companies to steal IP from US AI labs (Raphael Satter/Reuters)

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

Raphael Satter / Reuters: Diplomatic cable: US State Department has ordered a global push to bring attention to what it says are efforts by Chinese companies to steal IP from US AI labs — The U.S. Sta

EARL-BO: Reinforcement Learning for Multi-Step Lookahead, High-Dimensional Bayesian Optimization

Model ReleasesDGX agent

arXiv:2411.00171v2 Announce Type: replace Abstract: To avoid myopic behavior, multi-step lookahead Bayesian optimization (BO) algorithms consider the sequential nature of BO and have demonstrated prom

Fake or Real, Can Robots Tell? Evaluating VLM Robustness to Domain Shift in Single-View Robotic Scene Understanding

Model ReleasesDGX agent

arXiv:2506.19579v3 Announce Type: replace-cross Abstract: Robotic scene understanding increasingly relies on Vision-Language Models (VLMs) to generate natural language descriptions of the environment.

Flipping Against All Odds: Reducing LLM Coin Flip Bias via Verbalized Rejection Sampling

SafetyDGX agent

arXiv:2506.09998v2 Announce Type: replace-cross Abstract: Large language models (LLMs) can often accurately describe probability distributions using natural language, yet they still struggle to genera

For the last 72 hours since ml-intern launched we have had over 500+ autonomous AI research projects running on the Space at all times. Some…

Model ReleasesDGX agent

For the last 72 hours since ml-intern launched we have had over 500+ autonomous AI research projects running on the Space at all times. Some insane ones I saw: 1. A new AI paradigm from scratch — tryi

From Noise to Intent: Anchoring Generative VLA Policies with Residual Bridges

Local AiDGX agent

arXiv:2604.21391v1 Announce Type: cross Abstract: Bridging high-level semantic understanding with low-level physical control remains a persistent challenge in embodied intelligence, stemming from the

How VLAs (Really) Work In Open-World Environments

SafetyDGX agent

arXiv:2604.21192v1 Announce Type: cross Abstract: Vision-language-action models (VLAs) have been extensively used in robotics applications, achieving great success in various manipulation problems. Mo

Introducing DeepSeek V4 Pro, a long-context model with hybrid attention, three reasoning modes, and SOTA coding performance. AI natives can …

Model ReleasesDGX agent

Introducing DeepSeek V4 Pro, a long-context model with hybrid attention, three reasoning modes, and SOTA coding performance. AI natives can now use DeepSeek V4 Pro on Together AI and benefit from reli

Just merged a built-in skill for Google's DESIGN.md A skill that lets Hermes author, lint, diff, and export DESIGN.md files, giving it fluen…

ResearchDGX agent

Just merged a built-in skill for Google's DESIGN.md A skill that lets Hermes author, lint, diff, and export DESIGN.md files, giving it fluency in Google's new open-source visual-identity format the mo

KGiRAG: An Iterative GraphRAG Approach for Responding Sensemaking Queries

ResearchDGX agent

arXiv:2604.20859v1 Announce Type: cross Abstract: Recent literature highlights the potential of graph-based approaches within large language model (LLM) retrieval-augmented generation (RAG) pipelines

LiveSense: A Real-Time Wi-Fi Sensing Platform for Range-Doppler on COTS Laptop

Local AiDGX agent

arXiv:2603.06545v2 Announce Type: replace-cross Abstract: We present LiveSense - a cross-platform that transforms a commercial off-the-shelf (COTS) Wi-Fi Network Interface Card (NIC) on a laptop into

M-CARE: Standardized Clinical Case Reporting for AI Model Behavioral Disorders, with a 20-Case Atlas and Experimental Validation

ResearchDGX agent

arXiv:2604.20871v1 Announce Type: cross Abstract: We introduce M-CARE (Model Clinical Assessment and Reporting for Evaluation), a clinical case report framework for AI model behavioral disorders adapt

Machine Behavior in Relational Moral Dilemmas: Moral Rightness, Predicted Human Behavior, and Model Decisions

SafetyDGX agent

arXiv:2604.21871v1 Announce Type: new Abstract: Human moral judgment is context-dependent and modulated by interpersonal relationships. As large language models (LLMs) increasingly function as decisio

Measuring Opinion Bias and Sycophancy via LLM-based Coercion

Model ReleasesDGX agent

arXiv:2604.21564v1 Announce Type: new Abstract: Large language models increasingly shape the information people consume: they are embedded in search, consulted for professional advice, deployed as age

Mind the Prompt: Self-adaptive Generation of Task Plan Explanations via LLMs

SafetyDGX agent

arXiv:2604.21092v1 Announce Type: new Abstract: Integrating Large Language Models (LLMs) into complex software systems enables the generation of human-understandable explanations of opaque AI processe

Models are changing really fast, often in ways that will break your product and technical assumptions. There's no way around this – the only…

ToolsDGX agent

Models are changing really fast, often in ways that will break your product and technical assumptions. There's no way around this – the only way forward is to relentlessly iterate to make sure you're

More on CursorBench: http://cursor.com/blog/cursorbench

ToolsDGX agent

CursorBench is a benchmarking tool or evaluation framework developed by Cursor, likely designed to measure and compare the performance of AI code completion or development features. The post provides

New episode of The Information Bottleneck is out, this time with @liuzhuang1234 (Princeton). We talked about ConvNeXt and whether architectu…

SafetyDGX agent

New episode of The Information Bottleneck is out, this time with @liuzhuang1234 (Princeton). We talked about ConvNeXt and whether architecture still matters; dataset bias and what 'good data' actually

OptiVerse: A Comprehensive Benchmark towards Optimization Problem Solving

Model ReleasesDGX agent

arXiv:2604.21510v1 Announce Type: new Abstract: While Large Language Models (LLMs) demonstrate remarkable reasoning, complex optimization tasks remain challenging, requiring domain knowledge and robus

Preferences of a Voice-First Nation: Large-Scale Pairwise Evaluation and Preference Analysis for TTS in Indian Languages

ResearchDGX agent

arXiv:2604.21481v1 Announce Type: new Abstract: Crowdsourced pairwise evaluation has emerged as a scalable approach for assessing foundation models. However, applying it to Text to Speech(TTS) introdu

ReCAPA: Hierarchical Predictive Correction to Mitigate Cascading Failures

Local AiDGX agent

arXiv:2604.21232v1 Announce Type: new Abstract: Vision-Language-Action systems follow instructions to execute multi-step tasks in multimodal environments. Recent VLA approaches typically rely on post-

Strategic Polysemy in AI Discourse: A Philosophical Analysis of Language, Hype, and Power

SafetyDGX agent

arXiv:2604.21043v1 Announce Type: cross Abstract: This paper examines the strategic use of language in contemporary artificial intelligence (AI) discourse, focusing on the widespread adoption of metap

Thinking Like a Botanist: Challenging Multimodal Language Models with Intent-Driven Chain-of-Inquiry

Model ReleasesDGX agent

arXiv:2604.20983v1 Announce Type: cross Abstract: Vision evaluations are typically done through multi-step processes. In most contemporary fields, experts analyze images using structured, evidence-bas

This has been the most anticipated event in OSS and it doesn't disappoint. Congratulations to @deepseek_ai team. DSV4 Pro is now available o…

Model ReleasesDGX agent

This has been the most anticipated event in OSS and it doesn't disappoint. Congratulations to @deepseek_ai team. DSV4 Pro is now available on @togethercompute and we'll be adding a lot of capacity beh

Tracxn: global edtech funding fell from 16.7B in 2021 to 2.6B in 2025, while the number of startups launched dropped from 10,491 in 2021 to just 645 in 2025 (Ananya Bhattacharya/Rest of World)

IndustryDGX agent

Ananya Bhattacharya / Rest of World: Tracxn: global edtech funding fell from 16.7B in 2021 to 2.6B in 2025, while the number of startups launched dropped from 10,491 in 2021 to just 645 in 2025 — Vent

Transient Turn Injection: Exposing Stateless Multi-Turn Vulnerabilities in Large Language Models

Model ReleasesDGX agent

arXiv:2604.21860v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly integrated into sensitive workflows, raising the stakes for adversarial robustness and safety. This pape

Trustworthy Clinical Decision Support Using Meta-Predicates and Domain-Specific Languages

Model ReleasesDGX agent

arXiv:2604.21263v1 Announce Type: new Abstract: extbf{Background:} Regulatory frameworks for AI in healthcare, including the EU AI Act and FDA guidance on AI/ML-based medical devices, require clinical

When Bigger Isn't Better: A Comprehensive Fairness Evaluation of Political Bias in Multi-News Summarisation

SafetyDGX agent

arXiv:2604.21309v1 Announce Type: new Abstract: Multi-document news summarisation systems are increasingly adopted for their convenience in processing vast daily news content, making fairness across d

Why are all LLMs Obsessed with Japanese Culture? On the Hidden Cultural and Regional Biases of LLMs

SafetyDGX agent

arXiv:2604.21751v1 Announce Type: cross Abstract: LLMs have been showing limitations when it comes to cultural coverage and competence, and in some cases show regional biases such as amplifying Wester

23 Apr 2026

A Survey of Scaling in Large Language Model Reasoning

SafetyDGX agent

arXiv:2504.02181v2 Announce Type: replace Abstract: The rapid advancements in large Language models (LLMs) have significantly enhanced their reasoning capabilities, driven by various strategies such a

ALAS: Adaptive Long-Horizon Action Synthesis via Async-pathway Stream Disentanglement

ResearchDGX agent

arXiv:2604.20721v1 Announce Type: new Abstract: Long-Horizon (LH) tasks in Human-Scene Interaction (HSI) are complex multi-step tasks that require continuous planning, sequential decision-making, and

At #RenderCon2026, @EMostaque makes the case during How to Build the Holodeck that stories still matter, IP still matters, but the barriers …

TutorialsDGX agent

At #RenderCon2026, @EMostaque makes the case during How to Build the Holodeck that stories still matter, IP still matters, but the barriers to building entirely new experiences are dropping faster tha

Automatic Ontology Construction Using LLMs as an External Layer of Memory, Verification, and Planning for Hybrid Intelligent Systems

Model ReleasesDGX agent

arXiv:2604.20795v1 Announce Type: new Abstract: This paper presents a hybrid architecture for intelligent systems in which large language models (LLMs) are extended with an external ontological memory

Breaking the Assistant Mold: Modeling Behavioral Variation in LLM Based Procedural Character Generation

SafetyDGX agent

arXiv:2601.03396v2 Announce Type: replace Abstract: Procedural content generation has enabled vast virtual worlds through levels, maps, and quests, but large-scale character generation remains underex

CARLA-Air: Fly Drones Inside a CARLA World -- A Unified Infrastructure for Air-Ground Embodied Intelligence

Model ReleasesDGX agent

arXiv:2603.28032v2 Announce Type: replace-cross Abstract: The convergence of low-altitude economies, embodied intelligence, and air-ground cooperative systems creates growing demand for simulation inf

Fairness-Aware Multi-Group Target Detection in Online Discussion

SafetyDGX agent

arXiv:2407.11933v4 Announce Type: replace Abstract: Target-group detection is the task of detecting which group(s) a piece of content is ``directed at or about''. Applications include targeted marketi

From Fuzzy to Formal: Scaling Hospital Quality Improvement with AI

SafetyDGX agent

arXiv:2604.20055v1 Announce Type: new Abstract: Hospital Quality Improvement (QI) plays a critical role in optimizing healthcare delivery by translating high-level hospital goals into actionable solut

Generative Flow Networks for Model Adaptation in Digital Twins of Natural Systems

SafetyDGX agent

arXiv:2604.20707v1 Announce Type: new Abstract: Digital twins of natural systems must remain aligned with physical systems that evolve over time, are only partially observed, and are typically modeled

GPT-5.5 in Codex is a delight to work with: - Super sharp with responses - It understands intent better than any model - Great 'personality'…

Model ReleasesDGX agent

GPT-5.5 in Codex is a delight to work with: - Super sharp with responses - It understands intent better than any model - Great 'personality' - Gets lots of stuff done without pausing unnecessarily It

Graph2Counsel: Clinically Grounded Synthetic Counseling Dialogue Generation from Client Psychological Graphs

SafetyDGX agent

arXiv:2604.20382v1 Announce Type: new Abstract: Rising demand for mental health support has increased interest in using Large Language Models (LLMs) for counseling. However, adapting LLMs to this high

HiPO: Hierarchical Preference Optimization for Adaptive Reasoning in LLMs

Model ReleasesDGX agent

arXiv:2604.20140v1 Announce Type: new Abstract: Direct Preference Optimization (DPO) is an effective framework for aligning large language models with human preferences, but it struggles with complex

Instagram launches Instants, an app for sharing disappearing photos, in Italy and Spain, after rolling out an Instants feature in its main app in some regions (Sydney Bradley/Business Insider)

Model ReleasesDGX agent

Sydney Bradley / Business Insider: Instagram launches Instants, an app for sharing disappearing photos, in Italy and Spain, after rolling out an Instants feature in its main app in some regions — - In

I've been previewing this in Codex for a few weeks - it's very good! Had some great results from it having it run security reviews against c…

Model ReleasesDGX agent

I've been previewing this in Codex for a few weeks - it's very good! Had some great results from it having it run security reviews against code written using other models Introducing GPT-5.5 A new cla

KOCO-BENCH: Can Large Language Models Leverage Domain Knowledge in Software Development?

Model ReleasesDGX agent

arXiv:2601.13240v2 Announce Type: replace-cross Abstract: Large language models (LLMs) excel at general programming but struggle with domain-specific software development, necessitating domain special

looks like new Pareto frontiers across everything: - Context: 400K context in Codex and a 1M in API - API Pricing: 5/m input and 30/m outp…

Model ReleasesDGX agent

looks like new Pareto frontiers across everything: - Context: 400K context in Codex and a 1M in API - API Pricing: 5/m input and 30/m output tokens. - Codex improved its own inference speed 20% lol -

Maximum Entropy Semi-Supervised Inverse Reinforcement Learning

ResearchDGX agent

arXiv:2604.20074v1 Announce Type: new Abstract: A popular approach to apprenticeship learning (AL) is to formulate it as an inverse reinforcement learning (IRL) problem. The MaxEnt-IRL algorithm succe

MirrorBench: Evaluating Self-centric Intelligence in MLLMs by Introducing a Mirror

Model ReleasesDGX agent

arXiv:2604.14785v2 Announce Type: replace Abstract: Recent progress in Multimodal Large Language Models (MLLMs) has demonstrated remarkable advances in perception and reasoning, suggesting their poten

Model Capability Assessment and Safeguards for Biological Weaponization

Model ReleasesDGX agent

arXiv:2604.19811v1 Announce Type: cross Abstract: AI leaders and safety reports increasingly warn that advances in model reasoning may enable biological misuse, including by low-expertise users, while

Occupancy Reward Shaping: Improving Credit Assignment for Offline Goal-Conditioned Reinforcement Learning

SafetyDGX agent

arXiv:2604.20627v1 Announce Type: new Abstract: The temporal lag between actions and their long-term consequences makes credit assignment a challenge when learning goal-directed behaviors from data. G

Online Structure Learning and Planning for Autonomous Robot Navigation using Active Inference

Local AiDGX agent

arXiv:2510.09574v2 Announce Type: replace Abstract: Autonomous navigation in unfamiliar environments requires robots to simultaneously explore, localise, and plan under uncertainty, without relying on

ONOTE: Benchmarking Omnimodal Notation Processing for Expert-level Music Intelligence

Model ReleasesDGX agent

arXiv:2604.20719v1 Announce Type: cross Abstract: Omnimodal Notation Processing (ONP) represents a unique frontier for omnimodal AI due to the rigorous, multi-dimensional alignment required across aud

🚨 OpenAI just launched GPT-5.5. The OpenAI team was nice enough to give me early access over the last several weeks, and I just want to fla…

Model ReleasesDGX agent

🚨 OpenAI just launched GPT-5.5. The OpenAI team was nice enough to give me early access over the last several weeks, and I just want to flag: there is a certain class of models (one that we’re hitting

OThink-SRR1: Search, Refine and Reasoning with Reinforced Learning for Large Language Models

ResearchDGX agent

arXiv:2604.19766v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) expands the knowledge of Large Language Models (LLMs), yet current static retrieval methods struggle with complex

PR-CAD: Progressive Refinement for Unified Controllable and Faithful Text-to-CAD Generation with Large Language Models

Model ReleasesDGX agent

arXiv:2604.19773v1 Announce Type: cross Abstract: The construction of CAD models has traditionally relied on labor-intensive manual operations and specialized expertise. Recent advances in large langu

R2IF: Aligning Reasoning with Decisions via Composite Rewards for Interpretable LLM Function Calling

ResearchDGX agent

arXiv:2604.20316v1 Announce Type: new Abstract: Function calling empowers large language models (LLMs) to interface with external tools, yet existing RL-based approaches suffer from misalignment betwe

Relative Principals, Pluralistic Alignment, and the Structural Value Alignment Problem

SafetyDGX agent

arXiv:2604.20805v1 Announce Type: cross Abstract: The value alignment problem for artificial intelligence (AI) is often framed as a purely technical or normative challenge, sometimes focused on hypoth

SMARTER: A Data-efficient Framework to Improve Toxicity Detection with Explanation via Self-augmenting Large Language Models

Model ReleasesDGX agent

arXiv:2509.15174v3 Announce Type: replace-cross Abstract: WARNING: This paper contains examples of offensive materials. To address the proliferation of toxic content on social media, we introduce SMAR

The Ratchet Effect in Silico through Interaction-Driven Cumulative Intelligence in Large Language Models

Model ReleasesDGX agent

arXiv:2507.21166v2 Announce Type: replace-cross Abstract: Human intelligence scales through cumulative cultural evolution (CCE), a ratchet process in which innovations are retained against entropic dr

← Previous
1…284285286287288…294
Next →