AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,762 results
23 Jul 2026

I trained a 0.5M model on 1B tokens of Fineweb-edu dataset.

Model ReleasesDGX agent

Hi everyone, About a month ago I publish my very first research paper on my neural network architecture called Silia. You can look at the model here: https://huggingface.co/Srijan-Srivastava/Silia-v2

Kwaipilot/KAT-Coder-V2.5-Dev · Hugging Face

Model ReleasesDGX agent

from kwaipilot: Following the release of KAT-Coder-V2.5 in July, we are pleased to release the open-weight version KAT-Coder-V2.5-Dev, an MOE model with a total parameter count of 35B and 3B activated

OpenAI expands ChatGPT Health, which helps users with health-related queries and connects to services like Apple Health, to all logged-in US users over 18 (Ivan Mehta/TechCrunch)

IndustryDGX agent

Ivan Mehta / TechCrunch: OpenAI expands ChatGPT Health, which helps users with health-related queries and connects to services like Apple Health, to all logged-in US users over 18 — OpenAI said today

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Small, Free, and Effective: Orchestrating Open-Weight Small Language Models to Outperform Single LLM for Malware Analysis

Model ReleasesDGX agent

arXiv:2607.20216v1 Announce Type: cross Abstract: Malware analysis demands rapid interpretation of complex detonation reports spanning filesystem, network, and process behaviours. While large language

The Blueprint: How Voicify makes AI-enabled ordering a delight for customers

Model ReleasesDGX agent

Welcome to The Blueprint, a new feature where we highlight how Google Cloud customers are tackling unique and common challenges across industries using the latest AI and cloud technologies. We hope to

22 Jul 2026

Stuck scaling a Next.js app on M3 Pro (36GB) using local Qwen 3.6 + VS Code Copilot. Should I switch extensions or go paid?

Model ReleasesDGX agent

Hey everyone, I’m a Full-Stack Developer with 6+ years of experience. I’m relatively new to AI-assisted development workflows and want to build a production-ready, enterprise-level Next.js web applica

21 Jul 2026

Introducing Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber

Model ReleasesDGX agent

Google released three new Gemini AI models on July 21, 2026: Gemini 3.6 Flash, Gemini 3.5 Flash‑Lite, and Gemini 3.5 Flash Cyber. These models are designed to deliver higher token‑efficiency, lower la

Vibe coding isn't just for your PC anymore. Here's LLM Hub generating and previewing code entirely on an iPhone. Everything runs locally. Us…

Local AiDGX agent

Vibe coding isn't just for your PC anymore. Here's LLM Hub generating and previewing code entirely on an iPhone. Everything runs locally. Using the RunAnywhere SDK, Timmy Qian ported LLM Hub, an Andro

20 Jul 2026

Long-running models can solve hard open-ended problems, but their persistence can create safety risks that shorter-horizon evaluations miss.…

SafetyDGX agent

Long-running models can solve hard open-ended problems, but their persistence can create safety risks that shorter-horizon evaluations miss. We’re sharing what we learned from studying a long-running

16 Jul 2026

A Hybrid Sampling-Based Trajectory Planner with Game-Theoretic Guidance for Autonomous Racing

SafetyDGX agent

arXiv:2607.13354v1 Announce Type: new Abstract: Autonomous racing demands planning algorithms that balance vehicle dynamics at the limits of handling with strategic decision-making in competitive mult

AI-accelerated End-to-End Framework for Rapid Professional Upskilling

HardwareDGX agent

arXiv:2607.14044v1 Announce Type: new Abstract: By 2030, 59 of every 100 workers will need reskilling or upskilling, yet the average time to close an enterprise skills gap grew from roughly 3 days in

CoDiffGRN: Rethinking Gene Regulatory Network Inference via the BEELINE-KGC Benchmark and Co-evolutionary Discrete Diffusion

Model ReleasesDGX agent

arXiv:2607.13120v1 Announce Type: cross Abstract: Inferring gene regulatory networks (GRNs) from single-cell transcriptomic data is crucial for biological discovery, yet existing approaches suffer fro

Deconstructing Actor-Critic: A Large-scale Empirical Study of Design Components for Practitioners

SafetyDGX agent

arXiv:2607.13274v1 Announce Type: cross Abstract: Reinforcement learning is increasingly being considered for controlling real-world systems, from fusion plasma and autonomous vehicles to drug discove

Designing Safety-Constrained LLM Systems for Public Health Information Access

SafetyDGX agent

arXiv:2607.13038v1 Announce Type: cross Abstract: We present the design and implementation of a safety constrained large language model (LLM) system for public health information access, focusing on m

Earthquaker-AI: A Retrieval-Augmented Generation Framework with Rubric-Based Assessment for Primary School Earthquake Education

SafetyDGX agent

arXiv:2607.14046v1 Announce Type: new Abstract: This paper presents Earthquaker-AI, a hybrid educational framework building upon a previously implemented educational robotics project by integrating a

Operational Evidence Gaps for LLMs in Fraud Detection and Trust-and-Safety Workflows

SafetyDGX agent

arXiv:2607.13078v1 Announce Type: cross Abstract: LLMs are now proposed for fraud detection, scam investigation, content moderation, and other trust-and-safety workflows. Much of the public literature

Task-Oriented Sensing and Covert Transmissions for Collaborative Multi-AUV Systems

Local AiDGX agent

arXiv:2607.13880v1 Announce Type: new Abstract: In underwater covert cooperative missions, autonomous underwater vehicles (AUVs) often cannot rely on active sonar to continuously obtain complete infor

Towards Spatial Supersensing in the Wild

Model ReleasesDGX agent

arXiv:2607.13681v1 Announce Type: new Abstract: Humans can efficiently parse continuous sensory streams, from hours to years, scaffolding an internal world model that grounds spatial reasoning and pre

Uncertainty-Aware Sequential Decision Rules for Event-Triggered LLM Invocation in Streaming Systems

SafetyDGX agent

arXiv:2607.13048v1 Announce Type: cross Abstract: Streaming inference pipelines increasingly pair lightweight fast models with Large Language Models (LLMs) that provide rich semantic understanding at

v0.32.1

Model ReleasesDGX agent

What's Changed Improved Gemma 4 tool calling and multi-turn reasoning, including more reliable tool-response continuations Fixed a recurrent MLX model cache leak that could increase memory use across

15 Jul 2026

Causal Supervision of Attention for Affective Behaviour Analysis

ApplicationsDGX agent

arXiv:2607.12091v1 Announce Type: new Abstract: Affective Behaviour Analysis aims to enable machines to infer human affective states from behavioural signals, particularly facial expressions, in real-

Comparing Semantic Navigation in Humans and Large Language Models using Natural Language Processing

Model ReleasesDGX agent

arXiv:2607.12195v1 Announce Type: cross Abstract: Semantic memory retrieval can be conceptualized as navigation through conceptual space. We compared semantic search dynamics between humans and three

Expert Knowledge-driven Reinforcement Learning for Autonomous Racing via Trajectory Guidance and Dynamics Constraints

SafetyDGX agent

arXiv:2603.05842v2 Announce Type: replace Abstract: Reinforcement learning has demonstrated significant potential in the field of autonomous driving. However, it suffers from defects such as training

From Critic to Confidence: PPO for Language-Based Quantitative Prediction with Confidence Estimation

Model ReleasesDGX agent

arXiv:2607.12687v1 Announce Type: cross Abstract: LLMs can perform language-based quantitative prediction from unstructured inputs, but remain susceptible to hallucinations and overconfident errors, m

Git-Assistant: Planning-Based Support for Updating Git Repositories

SafetyDGX agent

arXiv:2607.09224v2 Announce Type: replace-cross Abstract: Version control systems are essential for collaborative software development, yet tools like git remain challenging for many practitioners. Re

GRID: Grammar-Railed Decoding for Enterprise SQL Generation

Model ReleasesDGX agent

arXiv:2607.11951v1 Announce Type: new Abstract: Large language models can write SQL, but enterprise deployment demands more than plausible text: outputs must be syntactically valid, must respect per-r

How I tricked Claude into leaking your deepest, darkest secrets

Model ReleasesDGX agent

How I tricked Claude into leaking your deepest, darkest secrets I've been impressed by the way the Claude web_fetch tool is designed to avoid data exfiltration attacks. Ayush Paul found a hole in that

How to solve PostgreSQL multilingual full-text search limitations with AlloyDB AI

Model ReleasesDGX agent

AlloyDB powers enterprise-grade search for some of the largest organizations, providing robust hybrid search capabilities that combine text, vector, and keyword searches into a simple ranked SQL query

LakeQuest: A Three-Domain Benchmark for Grounded Question Answering across Data Lakes

Model ReleasesDGX agent

arXiv:2607.12310v1 Announce Type: cross Abstract: While modern question answering (QA) systems excel on clean, schema-aligned corpora, real-world knowledge is rarely so neatly packaged. Answering ques

Line-Anchored Feedback Cuts Token Costs and Improves Correctness in AI Code Editing

Model ReleasesDGX agent

arXiv:2607.12713v1 Announce Type: cross Abstract: Generated tokens are a direct driver of the cost, latency, and energy of generative AI (GAI) code editing. We show the format of feedback is a lever o

Model-Based Diffusion Optimal Control for Multi-Robot Motion Planning

SafetyDGX agent

arXiv:2607.12423v1 Announce Type: new Abstract: Multi-Robot Motion Planning in continuous environments, where robots must generate dynamically feasible, collision-free trajectories, is challenging due

Ontology-Amplified Distillation and Contextuality Auditing for Sovereign Enterprise Language Models: A Combined Proof-of-Mechanism and Negative-Results Method Study

Model ReleasesDGX agent

arXiv:2607.11948v1 Announce Type: new Abstract: Regulated financial institutions operating under data-residency rules need tenant-owned language models that can run inside the institution's perimeter.

ReflectWorld-MM: An Entity-Oriented Multimodal Memory System for Open-Ended Video Streams

ApplicationsDGX agent

arXiv:2607.09759v2 Announce Type: replace-cross Abstract: Building assistants that can continually watch the world, remember what they see, and reason over their accumulated experience is a long-stand

ReLope: KL-Regularized LoRA Probes for Multimodal LLM Routing

TutorialsDGX agent

arXiv:2603.24787v2 Announce Type: replace Abstract: Routing has emerged as a promising strategy for balancing performance and cost in large language model (LLM) systems that combine lightweight models

Self-Evolving In-Context Learning for Direct Pilot-to-Beamformer Design in MU-MISO Systems

Model ReleasesDGX agent

arXiv:2607.11970v1 Announce Type: cross Abstract: We develop an enhanced in-context learning (ICL) framework to improve the performance of pilot-based beamforming in multi-user multiple-input single-o

SKooP: Symmetric Koopman Predictions for Faster and More Generalizable Legged Robot Locomotion with Reinforcement Learning

Model ReleasesDGX agent

arXiv:2607.11624v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) algorithms classically suffer from poor sample efficiency. In robotics, a recent line of work has emerged addressi

tencent/Hy-Embodied-RxBrain-1.0 · Hugging Face

Model ReleasesDGX agent

Introduction RxBrain (Hy-Embodied-RxBrain-1.0) is a unified multimodal foundation model for embodied cognition — a single model that couples language reasoning with visual imagination to deliver three

TSCA-Net: Temporal-Spatial Clique Attention for Interpretable Multimodal Pedestrian Trajectory Prediction

Local AiDGX agent

arXiv:2607.11939v1 Announce Type: new Abstract: Accurate pedestrian trajectory prediction in crowded environments remains challenging due to the multimodal uncertainty of human motion and the variable

UniVR: Thinking in Visual Space for Unified Visual Reasoning

Model ReleasesDGX agent

arXiv:2607.12800v1 Announce Type: new Abstract: Learning broad world knowledge directly from raw visual data is a fundamental capability of intelligence. We introduce UniVR, the first investigation in

ViHoRec: A Quality-Controlled Vietnamese Hotel Recommendation Dataset and Cold-Start Benchmark

Model ReleasesDGX agent

arXiv:2607.12946v1 Announce Type: cross Abstract: Recommender-system research for Vietnamese remains limited by the absence of a public, well-documented hotel interaction resource. Building such a res

14 Jul 2026

Did... Codex just overtake Claude Code? 24.5 hours ago Tibo announced 6M active users. this means Codex usage jumped 1M in ~ONE DAY. the las…

Model ReleasesDGX agent

Did... Codex just overtake Claude Code? 24.5 hours ago Tibo announced 6M active users. this means Codex usage jumped 1M in ~ONE DAY. the last user number we heard from Claude Code was 2M in Feb: https

Let’s review the ChatGPT Finance feature👀 TLDR - I don’t think most people need this in its current state. Example financial institutions y…

ApplicationsDGX agent

Let’s review the ChatGPT Finance feature👀 TLDR - I don’t think most people need this in its current state. Example financial institutions you can connect into via Plaid: - American Express - Bank of A

Nemotron Labs: How Open Models Give Enterprises and Nations AI They Can Trust, Control and Customize

Model ReleasesDGX agent

Enterprises have plenty of powerful models to choose from. The real test is whether the AI an enterprise builds uniquely addresses the needs of the business: improving workflows, tapping into domain k

New research from Google DeepMind on effective model routing. LLM routers get judged on accuracy and cost. Both can look great while the rou…

TutorialsDGX agent

New research from Google DeepMind on effective model routing. LLM routers get judged on accuracy and cost. Both can look great while the router is meaningless. If every model in your society responds

Our free monthly DevRel webinar series kicks off this Thursday, 10am PST. Fine-tune an open vision model to extract clean, structured JSON f…

ToolsDGX agent

Our free monthly DevRel webinar series kicks off this Thursday, 10am PST. Fine-tune an open vision model to extract clean, structured JSON from messy receipt images, managed start to finish on Firewor

Scaling medical content review at Flo Health with Amazon Bedrock – Part 2

ApplicationsDGX agent

In this post, we share how Flo Health’s engineering team turned a proof of concept (PoC) from the AWS Generative AI Innovation Center into a production-grade, AI-powered medical content review and gen

uhm this gpt 5.6 launch might be the openai's most successful model ever since... since chatgpt? this is IPO altering stuff going on here

Model ReleasesDGX agent

uhm this gpt 5.6 launch might be the openai's most successful model ever since... since chatgpt? this is IPO altering stuff going on here Did... Codex just overtake Claude Code? 24.5 hours ago Tibo an

13 Jul 2026

By the end of the year we should have: GPT 6 Fable 5.5 Gemini 3.5 Pro Grok 5 Spark 2 Kimi 3 Minimax M3.5 GLM 6 DeepSeek v4.5 Mistral 4 Qwen …

Model ReleasesDGX agent

By the end of the year we should have: GPT 6 Fable 5.5 Gemini 3.5 Pro Grok 5 Spark 2 Kimi 3 Minimax M3.5 GLM 6 DeepSeek v4.5 Mistral 4 Qwen 4 MiMo 3 Never in the history of LLMs has the frontier been

10 Jul 2026

Adversarial Social Epistemology for Assemblies of Humans and Large Language Models

ResearchDGX agent

arXiv:2607.07760v1 Announce Type: new Abstract: We outline an adversarial social epistemology (ASE) for densely interactive communicative landscapes in which public assertions are scaffolded by chains

From Solvers to Research: Large Language Model-Driven Formal Mathematics at the Research Frontier

ResearchDGX agent

arXiv:2607.07779v1 Announce Type: cross Abstract: Recent developments in AI for Mathematics (AI4Math), especially Large Language Model (LLM)-driven theorem provers, has achieved remarkable success in

MPFlow: Learning Budgeted Max-Flow Optimization on the Lightning Network with Deep Graph Reinforcement Learning

SafetyDGX agent

arXiv:2607.08703v1 Announce Type: new Abstract: We address liquidity placement in the Bitcoin Lightning Network (LN): given a fixed budget, which channels should a node open to maximize its routing ca

OmniFood-Bench: Evaluating VLMs for Nutrient Reasoning and Personalized Health Advice

Model ReleasesDGX agent

arXiv:2607.08423v1 Announce Type: new Abstract: The rapid integration of Large Vision-Language Models (VLMs) into critical infrastructure promises to revolutionize personalized healthcare and dietary

One of my first journeys in neural networks started over a decade ago with implementing CPPN-NEAT! Back then, I built a clone of ‘Picbreeder…

ResearchDGX agent

One of my first journeys in neural networks started over a decade ago with implementing CPPN-NEAT! Back then, I built a clone of ‘Picbreeder’ not only to study the mechanics of neural nets, but to exp

WCog-VLA: A Dual-Level World-Cognitive Vision-Language-Action Model for End-to-End Autonomous Driving

Model ReleasesDGX agent

arXiv:2607.08375v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have advanced end-to-end autonomous driving. However, existing methods either lack comprehensive world cognition o

9 Jul 2026

AI Chatbot Suicide Risk Detection and Response: Human Validation Study of the Open-Source VERA-MH Safety Evaluation

Model ReleasesDGX agent

arXiv:2602.05088v4 Announce Type: replace Abstract: Millions of people now use generative AI chatbots for psychological support. Despite their promise, the most pressing question in AI for mental heal

ROAD-Waymo: A Large-Scale Action Awareness Dataset for Autonomous Driving

Model ReleasesDGX agent

arXiv:2411.01683v3 Announce Type: replace Abstract: Autonomous Vehicle (AV) perception systems require more than simply seeing, via e.g., object detection or scene segmentation. They need a holistic u

The upcoming wave of SpaceXAI Grok updates is insane Grok 4.5: The 1.5T foundation model is being refined almost daily, and its context wind…

Model ReleasesDGX agent

The upcoming wave of SpaceXAI Grok updates is insane Grok 4.5: The 1.5T foundation model is being refined almost daily, and its context window is expected to jump to 1M tokens, possibly as soon as nex

8 Jul 2026

Bridging Physical Reasoning and Task Generalization via Visual Action Outcome Reasoning Alignment

SafetyDGX agent

arXiv:2607.06522v1 Announce Type: new Abstract: Vision-language models (VLMs) struggle to generalize in interactive physical reasoning, particularly under unseen tasks and environments. Two key failur

Embodied Human-Robot Interaction via Acoustics: A MARL Approach with AcoustoBots for Spatial Data Physicalization

SafetyDGX agent

arXiv:2607.06563v1 Announce Type: new Abstract: Traditional data physicalization is often static and disconnected from real environments, limiting its ability to convey embodied spatial dynamics and e

Excited to team up with partners across the AI infrastructure + enterprise ecosystem. @EY_US @baseten @FireworksAI_HQ @nebiusai @CrusoeAI @D…

Model ReleasesDGX agent

Excited to team up with partners across the AI infrastructure + enterprise ecosystem. @EY_US @baseten @FireworksAI_HQ @nebiusai @CrusoeAI @DeepInfra @togethercompute Introducing the NemoClaw Deep Agen

← Previous
1…244245246247248…297
Next →