AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
10,088 results
21 Apr 2026

VIDS: A Verified Imaging Dataset Standard for Medical AI

Model ReleasesDGX agent

arXiv:2604.17525v1 Announce Type: cross Abstract: Medical imaging AI development is fundamentally dependent on annotated datasets, yet no existing standard provides machine-enforceable validation acro

What is born of light is light

ResearchDGX agent

This post likely discusses how light or illumination (literal or metaphorical) generates or produces similar qualities, potentially referencing philosophical, scientific, or spiritual concepts about t

Where to Focus: Query-Modulated Multimodal Keyframe Selection for Long Video Understanding

Local AiDGX agent

arXiv:2604.17422v1 Announce Type: new Abstract: Long video understanding remains a formidable challenge for Multimodal Large Language Models (MLLMs) due to the prohibitive computational cost of proces

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Zero Networks launches AI Segmentation to govern autonomous AI agents

Model ReleasesDGX agent

Zero-trust security startup Zero Networks Ltd. today announced the launch of AI Segmentation, a set of platform capabilities designed to give enterprises identity-based control over autonomous artific

20 Apr 2026

AgentV-RL: Scaling Reward Modeling with Agentic Verifier

AgentsDGX agent

arXiv:2604.16004v1 Announce Type: cross Abstract: Verifiers have been demonstrated to enhance LLM reasoning via test-time scaling (TTS). Yet, they face significant challenges in complex domains. Error

COMPASS: Benchmarking Constrained Optimization in LLM Agents

Model ReleasesDGX agent

arXiv:2510.07043v2 Announce Type: replace Abstract: Human decision-making often involves constrained optimization. As LLM agents are deployed to assist with real-world tasks like travel planning, shop

dear god lol - new kimi model is a fucking beast. GPT 5.4 level coding, 76% cheaper than opus 4.7 and 100% open source / free to use, i mean…

Model ReleasesDGX agent

dear god lol - new kimi model is a fucking beast. GPT 5.4 level coding, 76% cheaper than opus 4.7 and 100% open source / free to use, i mean look at this: > kimi k2.6 can code continuously for 12 hour

FACTS: Table Summarization via Offline Template Generation with Agentic Workflows

AgentsDGX agent

arXiv:2510.13920v2 Announce Type: replace Abstract: Query-focused table summarization requires generating natural language summaries of tabular data conditioned on a user query, enabling users to acce

Find, Fix, Reason: Context Repair for Video Reasoning

SafetyDGX agent

arXiv:2604.16243v1 Announce Type: new Abstract: Reinforcement learning has advanced video reasoning in large multi-modal models, yet dominant pipelines either rely on on-policy self-exploration, which

Get hands on with agents, vibe coding and more at Data+ AI Summit

IndustryDGX agent

The Data+ AI Summit features hands-on workshops and sessions covering practical applications of AI agents, vibe coding (a Databricks term for intuitive, conversational AI development), and other emerg

I've been a K2.5 superfan since it came out. These new numbers for the next version look incredible. You gotta love competition!

Model ReleasesDGX agent

I've been a K2.5 superfan since it came out. These new numbers for the next version look incredible. You gotta love competition! Meet Kimi K2.6: Advancing Open-Source Coding 🔹Open-source SOTA on HLE w

Kimi K2.6 is now available in Hermes Agent. Simply run `hermes update` and use `hermes model` to select a compatible provider hosting the mo…

AgentsDGX agent

Kimi K2.6 is now available in Hermes Agent. Simply run `hermes update` and use `hermes model` to select a compatible provider hosting the model! Meet Kimi K2.6: Advancing Open-Source Coding 🔹Open-sour

Kimi K2.6 raises the bar for open-source models. 🦙 available on Ollama's cloud! Try it with OpenClaw: ollama launch openclaw --model kimi-k…

Model ReleasesDGX agent

Kimi K2.6 raises the bar for open-source models. 🦙 available on Ollama's cloud! Try it with OpenClaw: ollama launch openclaw --model kimi-k2.6:cloud Try it with Hermes Agent: ollama launch hermes --mo

KRONE: Scalable LLM-Augmented Log Anomaly Detection via Hierarchical Abstraction

Local AiDGX agent

arXiv:2602.07303v3 Announce Type: replace-cross Abstract: Log anomaly detection is crucial for uncovering system failures and security risks. Although logs originate from nested component executions w

LaMSUM: Amplifying Voices Against Harassment through LLM Guided Extractive Summarization of User Incident Reports

Model ReleasesDGX agent

arXiv:2406.15809v5 Announce Type: replace Abstract: Citizen reporting platforms help the public and authorities stay informed about sexual harassment incidents. However, the high volume of data shared

life when you discover an open-source model that runs 300 parallel agents, executes for 12+ hours straight, beats GPT-5.4 and opus 4.6 on mu…

Model ReleasesDGX agent

life when you discover an open-source model that runs 300 parallel agents, executes for 12+ hours straight, beats GPT-5.4 and opus 4.6 on multiple benchmarks... and the weights are on huggingface Medi

LinuxArena: A Control Setting for AI Agents in Live Production Software Environments

Model ReleasesDGX agent

arXiv:2604.15384v1 Announce Type: cross Abstract: We introduce LinuxArena, a control setting in which agents operate directly on live, multi-service production environments. LinuxArena contains 20 env

MemEvoBench: Benchmarking Memory MisEvolution in LLM Agents

Model ReleasesDGX agent

arXiv:2604.15774v1 Announce Type: new Abstract: Equipping Large Language Models (LLMs) with persistent memory enhances interaction continuity and personalization but introduces new safety risks. Speci

PolicyBank: Evolving Policy Understanding for LLM Agents

Model ReleasesDGX agent

arXiv:2604.15505v1 Announce Type: cross Abstract: LLM agents operating under organizational policies must comply with authorization constraints typically specified in natural language. In practice, su

Security Threat Modeling for Emerging AI-Agent Protocols: A Comparative Analysis of MCP, A2A, Agora, and ANP

AgentsDGX agent

arXiv:2602.11327v2 Announce Type: replace-cross Abstract: The rapid development of the AI agent communication protocols, including the Model Context Protocol (MCP), Agent2Agent (A2A), Agora, and Agent

SENSE: Stereo OpEN Vocabulary SEmantic Segmentation

AgentsDGX agent

arXiv:2604.15946v1 Announce Type: new Abstract: Open-vocabulary semantic segmentation enables models to segment objects or image regions beyond fixed class sets, offering flexibility in dynamic enviro

VeriGraph: Scene Graphs for Execution Verifiable Robot Planning

AgentsDGX agent

arXiv:2411.10446v3 Announce Type: replace-cross Abstract: Recent progress in vision-language models (VLMs) has opened new possibilities for robot task planning, but these models often produce incorrec

19 Apr 2026

🆕 The Future of MCP https://www.youtube.com/watch?v=v3Fr2JR47KA London is the home of MCP, which is just a little over a year old and now t…

AgentsDGX agent

🆕 The Future of MCP https://www.youtube.com/watch?v=v3Fr2JR47KA London is the home of MCP, which is just a little over a year old and now the most successful AI integration protocol ever! @dsp_'s keyn

This is HUGE. Ollama now supports Hermes AI Agents. A fully local self-improving AI Agent that runs for you 24/7 for free.

Local AiDGX agent

Ollama has added support for Hermes AI Agents, enabling users to run fully local, self-improving AI agents on their machines continuously at no cost. This development allows autonomous agents to opera

18 Apr 2026

Hermes Agent 🤝 Ollama

Local AiDGX agent

Nous Research announced integration or collaboration between Hermes (their AI agent framework) and Ollama (a local LLM runtime platform), enabling users to run Hermes agents with locally-hosted langua

Nice paper from Google. And a great application of AI agents. Wearables capture a staggering amount of physiological signals every day. CoDa…

AgentsDGX agent

Nice paper from Google. And a great application of AI agents. Wearables capture a staggering amount of physiological signals every day. CoDaS is an AI co-data-scientist that turns raw wearable sensor

@openclaw And of course @Ollama for the local model-serving engine. 🦙

Local AiDGX agent

Ollama is a local model-serving engine that enables users to run large language models on their own hardware without relying on cloud services. The post appears to highlight Ollama's integration with

@otium33 hermes is like seeing color for the first time

ResearchDGX agent

This post likely expresses enthusiasm about Hermes, an AI model, comparing the experience of using it to a profound sensory revelation. The metaphor suggests that Hermes represents a significant quali

so hes building a harness

AgentsDGX agent

so hes building a harness Now launching GBrain v0.11 with Minions I got sick of OpenClaw's subagents timing out and not getting things done So I built a queue/jobs system that uses GBrain's Postgres/P

17 Apr 2026

Beyond Translation: Evaluating Mathematical Reasoning Capabilities of LLMs in Sinhala and Tamil

ResearchDGX agent

arXiv:2602.14517v2 Announce Type: replace Abstract: Large language models (LLMs) have achieved strong results in mathematical reasoning, and are increasingly deployed as tutoring and learning support

Does RL Expand the Capability Boundary of LLM Agents? A PASS@(k,T) Analysis

AgentsDGX agent

arXiv:2604.14877v1 Announce Type: new Abstract: Does reinforcement learning genuinely expand what LLM agents can do, or merely make them more reliable? For static reasoning, recent work answers the se

EchoAgent: Towards Reliable Echocardiography Interpretation with 'Eyes','Hands' and 'Minds'

AgentsDGX agent

arXiv:2604.05541v2 Announce Type: replace Abstract: Reliable interpretation of echocardiography (Echo) is crucial for assessing cardiac function, which demands clinicians to synchronously orchestrate

Fastest way to deploy deepagents gets even more powerful

ApplicationsDGX agent

Fastest way to deploy deepagents gets even more powerful we just shipped support for subagents with `deepagents deploy`! add an agents/ dir to your project with an AGENTS.md per specialized subagent.

From Reactive to Proactive: Assessing the Proactivity of Voice Agents via ProVoice-Bench

ResearchDGX agent

arXiv:2604.15037v1 Announce Type: cross Abstract: Recent advancements in LLM agents are gradually shifting from reactive, text-based paradigms toward proactive, multimodal interaction. However, existi

GraphScout: Empowering Large Language Models with Intrinsic Exploration Ability for Agentic Graph Reasoning

Model ReleasesDGX agent

arXiv:2603.01410v2 Announce Type: replace Abstract: Knowledge graphs provide structured and reliable information for many real-world applications, motivating increasing interest in combining large lan

Hermes creates skills from experience, improves them during use, nudges itself to persist knowledge, searches its own past conversations, an…

Local AiDGX agent

Hermes creates skills from experience, improves them during use, nudges itself to persist knowledge, searches its own past conversations, and builds a deepening model of who you are across sessions. D

HUOZIIME: An On-Device LLM-enhanced Input Method for Deep Personalization

Local AiDGX agent

arXiv:2604.14159v1 Announce Type: new Abstract: Mobile input method editors (IMEs) are the primary interface for text input, yet they remain constrained to manual typing and struggle to produce person

Interpretable and Explainable Surrogate Modeling for Simulations: A State-of-the-Art Survey and Perspectives on Explainable AI for Decision-Making

AgentsDGX agent

arXiv:2604.14240v1 Announce Type: cross Abstract: The simulation of complex systems increasingly relies on sophisticated but fundamentally opaque computational black-box simulators. Surrogate models p

Introducing the Agent Readiness score. Is your site agent-ready?

AgentsDGX agent

The Agent Readiness score can help site owners understand how well their websites support AI agents. Here we explore new standards, share Radar data, and detail how we made Cloudflare’s docs the most

Man with @ihackedthegovernment Instagram account tells judge, “I made a mistake'

IndustryDGX agent

Nicholas Moore, who pleaded guilty to hacking the U.S. Supreme Court's electronic document filing system dozens of times over several months, was sentenced on Friday to a year of probation. Moore also

MAS-Bench: A Unified Benchmark for Shortcut-Augmented Hybrid Mobile GUI Agents

Model ReleasesDGX agent

arXiv:2509.06477v2 Announce Type: replace Abstract: Shortcuts such as APIs and deep-links have emerged as efficient complements to flexible GUI operations, fostering a promising hybrid paradigm for ML

NanoClaw partners with Vercel to deliver one-click approvals for AI agents working on sensitive tasks

IndustryDGX agent

NanoCo, the startup behind NanoClaw, a fast-growing alternative to OpenAI Group PBC’s OpenClaw project, said today it’s teaming up with Vercel Inc. and OneCLI to try to fix the “trust problem” holding

ollama launch hermes Ollama 0.21 includes supports Hermes Agent, the self-improving AI agent built by @NousResearch.

Local AiDGX agent

Ollama version 0.21 added support for Hermes Agent, a self-improving AI agent developed by Nous Research. This release enables users to run Hermes Agent through the Ollama platform, expanding its capa

Recent advances push Big Tech closer to the Q-Day danger zone

IndustryDGX agent

Q-Day—the moment when quantum computers can break widely used cryptography—may be approaching faster than expected due to recent algorithmic and hardware advances. Advances in algorithms and design ar

Rethinking AI Hardware: A Three-Layer Cognitive Architecture for Autonomous Agents

Local AiDGX agent

arXiv:2604.13757v1 Announce Type: new Abstract: The next generation of autonomous AI systems will be constrained not only by model capability, but by how intelligence is structured across heterogeneou

Rethinking Patient Education as Multi-turn Multi-modal Interaction

Model ReleasesDGX agent

arXiv:2604.14656v1 Announce Type: cross Abstract: Most medical multimodal benchmarks focus on static tasks such as image question answering, report generation, and plain-language rewriting. Patient ed

SpaceMind: A Modular and Self-Evolving Embodied Vision-Language Agent Framework for Autonomous On-orbit Servicing

AgentsDGX agent

arXiv:2604.14399v1 Announce Type: new Abstract: Autonomous on-orbit servicing demands embodied agents that perceive through visual sensors, reason about 3D spatial situations, and execute multi-phase

The Cost of Language: Centroid Erasure Exposes and Exploits Modal Competition in Multimodal Language Models

Local AiDGX agent

arXiv:2604.14363v1 Announce Type: new Abstract: Multimodal language models systematically underperform on visual perception tasks, yet the structure underlying this failure remains poorly understood.

Unweight: how we compressed an LLM 22% without sacrificing quality

HardwareDGX agent

Running LLMs across Cloudflare’s network requires us to be smarter and more efficient about GPU memory bandwidth. That’s why we developed Unweight, a lossless inference-time compression system that ac

V-Reflection: Transforming MLLMs from Passive Observers to Active Interrogators

Local AiDGX agent

arXiv:2604.03307v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable success, yet they remain prone to perception-related hallucinations in fine-graine

we just shipped support for subagents with `deepagents deploy`! add an agents/ dir to your project with an AGENTS.md per specialized subagen…

ApplicationsDGX agent

we just shipped support for subagents with `deepagents deploy`! add an agents/ dir to your project with an AGENTS.md per specialized subagent. subagents are great for task delegation with isolated/opt

what should we ship next in `deepagents deploy`?

ApplicationsDGX agent

what should we ship next in `deepagents deploy`? we just shipped support for subagents with `deepagents deploy`! add an agents/ dir to your project with an AGENTS.md per specialized subagent. subagent

🇺🇸 xAI just dropped a real-time speech-to-text model built for voice apps: high limits, multi-language support, the works. Priced at just …

Model ReleasesDGX agent

🇺🇸 xAI just dropped a real-time speech-to-text model built for voice apps: high limits, multi-language support, the works. Priced at just 0.10–0.20/hour, crushing most competitors. Quiet release, loud

16 Apr 2026

A closer look at how large language models trust humans: patterns and biases

ResearchDGX agent

arXiv:2504.15801v2 Announce Type: replace Abstract: As large language models (LLMs) and LLM-based agents increasingly interact with humans in decision-making contexts, understanding the trust dynamics

A Comprehensive Survey on Network Traffic Synthesis: From Statistical Models to Deep Learning

ApplicationsDGX agent

arXiv:2507.01976v2 Announce Type: replace-cross Abstract: Synthetic network traffic generation has emerged as a promising alternative for various data-driven applications in the networking domain. It

AgentSPEX: An Agent SPecification and EXecution Language

AgentsDGX agent

arXiv:2604.13346v1 Announce Type: new Abstract: Language-model agent systems commonly rely on reactive prompting, in which a single instruction guides the model through an open-ended sequence of reaso

Antioch prepares to accelerate simulated testing for autonomous robots after raising $8.5M

AgentsDGX agent

Antioch Inc., a developer of cloud-based simulation software for artificial intelligence-enabled robots, has raised 8.5 million in funding to accelerate the development of more autonomous systems outs

Canva unveils Canva AI 2.0, recasting its platform as an agentic system for work

AgentsDGX agent

Visual communication platform provider Canva Pty Ltd. today unveiled Canva AI 2.0, an overhaul of its platform that offers a conversational, agentic system aimed at becoming the place where teams star

Impinj boosts edge computing power in updated R700 RAIN RFID reader

AgentsDGX agent

Radio-frequency identification devices and software company Impinj Inc. today announced an upgrade of its R700 RAIN RFID reader with a more powerful processor and expanded memory to help enterprises b

InfiniteScienceGym: An Unbounded, Procedurally-Generated Benchmark for Scientific Analysis

Model ReleasesDGX agent

arXiv:2604.13201v1 Announce Type: new Abstract: Large language models are emerging as scientific assistants, but evaluating their ability to reason from empirical data remains challenging. Benchmarks

← Previous
1…9899100101102…169
Next →