AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
9,953 results
10 Apr 2026

Fighting AI with AI: AI-Agent Augmented DNS Blocking of LLM Services during Student Evaluations

Model ReleasesDGX agent

arXiv:2604.02360v1 Announce Type: cross Abstract: The transformative potential of large language models (LLMs) in education, such as improving accessibility and personalized learning, is being eclipse

Financial services

ApplicationsDGX agent

The OpenAI Academy Financial Services page covers how ChatGPT can be applied across core financial workflows. Finance teams can leverage ChatGPT to accelerate financial analysis, automate routine d...

Flemme: A Flexible and Modular Learning Platform for Medical Images

ResearchDGX agent

arXiv:2408.09369v3 Announce Type: replace-cross Abstract: As the rapid development of computer vision and the emergence of powerful network backbones and architectures, the application of deep learnin

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Floating or Suggesting Ideas? A Large-Scale Contrastive Analysis of Metaphorical and Literal Verb-Object Constructions

ResearchDGX agent

arXiv:2604.08275v1 Announce Type: new Abstract: Metaphor pervades everyday language, allowing speakers to express abstract concepts via concrete domains. While prior work has studied metaphors cogniti

From Fragments to Facts: A Curriculum-Driven DPO Approach for Generating Hindi News Veracity Explanations

Model ReleasesDGX agent

arXiv:2507.05179v4 Announce Type: replace Abstract: In an era of rampant misinformation, generating reliable news explanations is vital, especially for under-represented languages like Hindi. Lacking

Garry @Kasparov63 retired from competitive chess over twenty years ago, and most or all of his tournament games are presumably publicly avai…

SafetyDGX agent

Garry @Kasparov63 retired from competitive chess over twenty years ago, and most or all of his tournament games are presumably publicly available - and yet I bet he still could crush any LLM that didn

Geometric Properties of the Voronoi Tessellation in Latent Semantic Manifolds of Large Language Models

SafetyDGX agent

arXiv:2604.06767v1 Announce Type: new Abstract: Language models operate on discrete tokens but compute in continuous vector spaces, inducing a Voronoi tessellation over the representation manifold. We

Getting started with ChatGPT

TutorialsDGX agent

OpenAI Academy's 'Getting Started with ChatGPT' tutorial introduces users to ChatGPT as a conversational AI application built on large language models, teaching core concepts such as how to write e...

glm 5.1 is doing well

Local AiDGX agent

GLM-5.1 is Z.ai's next-generation flagship model for agentic engineering, built on a 754-billion parameter Mixture-of-Experts architecture with 40 billion active parameters per token, a 200,000-tok...

Grok 4.20 hitting 83% on non-hallucination. Values truth. Claude ~74%. Others sitting in the 60s… or way lower. Less guessing. More honesty …

Model ReleasesDGX agent

Grok 4.20 hitting 83% on non-hallucination. Values truth. Claude ~74%. Others sitting in the 60s… or way lower. Less guessing. More honesty when it doesn’t know. That’s a different kind of intelligenc

How do I know if an AI model could work locally on my computer?

Local AiDGX agent

To determine if an AI model can run locally on your computer, the key factors are RAM, storage, and GPU availability: a modern PC with at least 8GB of RAM and a dedicated GPU is generally sufficien...

How FLORA shipped a creative agent on Vercel's AI stack

AgentsDGX agent

FLORA built FAUNA, a long-running creative agent that turns ideas into visual workflows on a digital canvas, allowing designers to describe goals like campaign visuals or moodboards and have the ag...

@hwchase17 middleware was the right abstraction for it too. way more adoptable than asking everyone to restructure their agent setup

AgentsDGX agent

LangChain's Middleware abstraction, introduced by Harrison Chase (@hwchase17) in LangChain 1.0 Alpha, addresses context engineering in AI agents by providing clean `before_model`, `after_model`, an...

I built AmicoScript: A local-first Whisper UI with Speaker Diarization and Ollama integration for summaries.

Local AiDGX agent

AmicoScript is a local-first, FastAPI-based application that combines OpenAI's Whisper for accurate speech-to-text transcription with Speaker Diarization to identify and separate different speakers...

Improving Robustness In Sparse Autoencoders via Masked Regularization

ResearchDGX agent

arXiv:2604.06495v1 Announce Type: cross Abstract: Sparse autoencoders (SAEs) are widely used in mechanistic interpretability to project LLM activations onto sparse latent spaces. However, sparsity alo

In a world where writing code to build websites and apps is trivial (thank you Lovable, Cursor, Claude,...), the real differentiation for yo…

Model ReleasesDGX agent

In a world where writing code to build websites and apps is trivial (thank you Lovable, Cursor, Claude,...), the real differentiation for you and your company (and what makes you successful) will be h

Is the ASUS ROG Flow Z13 with 128GB of Unified Memory (AMD Strix Halo) a good option to run large LLMs (70B+)?

Local AiDGX agent

The ASUS ROG Flow Z13 (2025) with AMD Ryzen AI Max+ 395 (Strix Halo) and 128GB of unified LPDDR5X memory is a capable portable option for running large LLMs locally, with ASUS officially stating it...

KITE: Keyframe-Indexed Tokenized Evidence for VLM-Based Robot Failure Analysis

Model ReleasesDGX agent

arXiv:2604.07034v1 Announce Type: cross Abstract: We present KITE, a training-free, keyframe-anchored, layout-grounded front-end that converts long robot-execution videos into compact, interpretable t

Learning Geometry-Aware Nonprehensile Pushing and Pulling with Dexterous Hands

ApplicationsDGX agent

arXiv:2509.18455v4 Announce Type: replace Abstract: Nonprehensile manipulation, such as pushing and pulling, enables robots to move, align, or reposition objects that may be difficult to grasp due to

Looking Beyond the Obvious: A Survey on Abstract Concept Recognition for Video Understanding

ResearchDGX agent

arXiv:2508.20765v2 Announce Type: replace-cross Abstract: The automatic understanding of video content is advancing rapidly. Empowered by deeper neural networks and large datasets, machines are increa

MCP is very much alive! With co-creator of MCP @dsp_ at AIE London

AgentsDGX agent

The referenced tweet (status ID 2042696610614849548) is not publicly accessible via search, and the X.com URL requires JavaScript/login to render. Based on the available context from related search...

middleware is underrated

AgentsDGX agent

I wasn't able to retrieve the content of that specific tweet or URL from the search results. X (Twitter) posts are generally not indexed in a way that allows direct retrieval of their content, and ...

Mina: A Multilingual LLM-Powered Legal Assistant Agent for Bangladesh for Empowering Access to Justice

AgentsDGX agent

arXiv:2511.08605v3 Announce Type: replace Abstract: Bangladesh's low-income population faces major barriers to affordable legal advice due to complex legal language, procedural opacity, and high costs

Multi-modal user interface control detection using cross-attention

ResearchDGX agent

arXiv:2604.06934v1 Announce Type: cross Abstract: Detecting user interface (UI) controls from software screenshots is a critical task for automated testing, accessibility, and software analytics, yet

MVOS_HSI: A Python Library for Preprocessing Agricultural Crop Hyperspectral Data

ResearchDGX agent

arXiv:2604.07656v1 Announce Type: cross Abstract: Hyperspectral imaging (HSI) allows researchers to study plant traits non-destructively. By capturing hundreds of narrow spectral bands per pixel, it r

Open-Ended Instruction Realization with LLM-Enabled Multi-Planner Scheduling in Autonomous Vehicles

Model ReleasesDGX agent

arXiv:2604.08031v1 Announce Type: cross Abstract: Most Human-Machine Interaction (HMI) research overlooks the maneuvering needs of passengers in autonomous driving (AD). Natural language offers an int

Open Harness, separated from model providers is a critical architectural pattern.

AgentsDGX agent

An **Open Harness** is a unified architectural layer that sits between AI agents and model providers, abstracting away provider-specific APIs and patterns. Because every AI agent harness has its o...

OpenClaude com Ollama Cloud

Local AiDGX agent

OpenClaude is an open-source coding-agent CLI, forked from the Claude Code source, that adds an OpenAI-compatible provider shim enabling use of GPT-4o, DeepSeek, Gemini, Ollama local models, and 20...

Possible memory leak in Ollama when using Claude Code?

Model ReleasesDGX agent

Users in the r/ollama community have reported a possible memory leak occurring in Ollama when it is used as a backend with Claude Code, with Ollama runner processes not always being properly termin...

Qualixar OS: A Universal Operating System for AI Agent Orchestration

SafetyDGX agent

arXiv:2604.06392v1 Announce Type: new Abstract: We present Qualixar OS, the first application-layer operating system for universal AI agent orchestration. Unlike kernel-level approaches (AIOS) or sing

Reasoning Fails Where Step Flow Breaks

ResearchDGX agent

arXiv:2604.06695v1 Announce Type: new Abstract: Large reasoning models (LRMs) that generate long chains of thought now perform well on multi-step math, science, and coding tasks. However, their behavi

Reasoning Within the Mind: Dynamic Multimodal Interleaving in Latent Space

SafetyDGX agent

arXiv:2512.12623v3 Announce Type: replace-cross Abstract: Recent advancements in Multimodal Large Language Models (MLLMs) have significantly enhanced cross-modal understanding and reasoning by incorpo

Responsible and safe use of AI

SafetyDGX agent

The OpenAI Academy page on 'Responsible and Safe Use of AI' is a guidance resource focused on best practices for using ChatGPT responsibly in professional and personal settings. It emphasizes that ...

Riemann-Bench: A Benchmark for Moonshot Mathematics

Model ReleasesDGX agent

arXiv:2604.06802v1 Announce Type: new Abstract: Recent AI systems have achieved gold-medal-level performance on the International Mathematical Olympiad, demonstrating remarkable proficiency at competi

RoboAgent: Chaining Basic Capabilities for Embodied Task Planning

SafetyDGX agent

arXiv:2604.07774v1 Announce Type: cross Abstract: This paper focuses on embodied task planning, where an agent acquires visual observations from the environment and executes atomic actions to accompli

SALLIE: Safeguarding Against Latent Language & Image Exploits

Model ReleasesDGX agent

arXiv:2604.06247v1 Announce Type: cross Abstract: Large Language Models (LLMs) and Vision-Language Models (VLMs) remain highly vulnerable to textual and visual jailbreaks, as well as prompt injections

Sampling-Aware 3D Spatial Analysis in Multiplexed Imaging

ResearchDGX agent

arXiv:2604.07890v1 Announce Type: new Abstract: Highly multiplexed microscopy enables rich spatial characterization of tissues at single-cell resolution, yet most analyses rely on two-dimensional sect

sciwrite-lint: Verification Infrastructure for the Age of Science Vibe-Writing

Model ReleasesDGX agent

arXiv:2604.08501v1 Announce Type: cross Abstract: Science currently offers two options for quality assurance, both inadequate. Journal gatekeeping claims to verify both integrity and contribution, but

SealQA: Raising the Bar for Reasoning in Search-Augmented Language Models

Model ReleasesDGX agent

arXiv:2506.01062v4 Announce Type: replace Abstract: We introduce SealQA, a new challenge benchmark for evaluating SEarch-Augmented Language models on fact-seeking questions where web search yields con

SkillClaw: Let Skills Evolve Collectively with Agentic Evolver

AgentsDGX agent

arXiv:2604.08377v1 Announce Type: cross Abstract: Large language model (LLM) agents such as OpenClaw rely on reusable skills to perform complex tasks, yet these skills remain largely static after depl

Spectral Edge Dynamics Reveal Functional Modes of Learning

Model ReleasesDGX agent

arXiv:2604.06256v1 Announce Type: cross Abstract: Training dynamics during grokking concentrate along a small number of dominant update directions -- the spectral edge -- which reliably distinguishes

Sub-agent Model Selection — Different Tasks, Different Models Your main agent runs Qwen3.6-Plus for quality. But not every subtask needs a f…

AgentsDGX agent

Sub-agent Model Selection — Different Tasks, Different Models Your main agent runs Qwen3.6-Plus for quality. But not every subtask needs a flagship model. Now sub-agents can use a different model. Cre

TGIF! Here are some of our favorite updates from the past week: — Notebooks in @GeminiApp, an integration with @NotebookLM that enables you …

Model ReleasesDGX agent

TGIF! Here are some of our favorite updates from the past week: — Notebooks in @GeminiApp, an integration with @NotebookLM that enables you to retrieve context from your private notebooks or convert y

The AI Skills Shift: Mapping Skill Obsolescence, Emergence, and Transition Pathways in the LLM Era

Model ReleasesDGX agent

arXiv:2604.06906v1 Announce Type: cross Abstract: As Large Language Models reshape the global labor market, policymakers and workers need empirical data on which occupational skills may be most suscep

The Depth Ceiling: On the Limits of Large Language Models in Discovering Latent Planning

Model ReleasesDGX agent

arXiv:2604.06427v1 Announce Type: cross Abstract: The viability of chain-of-thought (CoT) monitoring hinges on models being unable to reason effectively in their latent representations. Yet little is

The reaction people are having to AIs that can find bugs in code is fascinating. Finally, we have the capacity to fix the crisis in computer…

SafetyDGX agent

The reaction people are having to AIs that can find bugs in code is fascinating. Finally, we have the capacity to fix the crisis in computer security we’ve had for decades, and everyone is treating it

The Unreasonable Effectiveness of Data for Recommender Systems

ResearchDGX agent

arXiv:2604.06420v2 Announce Type: cross Abstract: In recommender systems, collecting, storing, and processing large-scale interaction data is increasingly costly in terms of time, energy, and computat

tldr > evals are the new training data. instead of updating weights, you're updating the agent harness > problem is agents are famous cheate…

AgentsDGX agent

tldr > evals are the new training data. instead of updating weights, you're updating the agent harness > problem is agents are famous cheaters. they will reward-hack your evals and overfit just to mak

Training Data Size Sensitivity in Unsupervised Rhyme Recognition

Model ReleasesDGX agent

arXiv:2604.08156v1 Announce Type: new Abstract: Rhyme is deceptively intuitive: what is or is not a rhyme is constructed historically, scholars struggle with rhyme classification, and people disagree

Training-free Spatially Grounded Geometric Shape Encoding (Technical Report)

ResearchDGX agent

arXiv:2604.07522v1 Announce Type: new Abstract: Positional encoding has become the de facto standard for grounding deep neural networks on discrete point-wise positions, and it has achieved remarkable

Ultraplan uses roughly the same number of tokens (and subscription rate limits) as plan mode. See the docs for more: http://docs.claude.com/…

Model ReleasesDGX agent

Ultraplan, a planning feature in Claude, consumes approximately the same number of tokens and counts against subscription rate limits similarly to standard plan mode. Users should be aware that using

Understanding Structured Financial Data with LLMs: A Case Study on Fraud Detection

ApplicationsDGX agent

arXiv:2512.13040v2 Announce Type: replace-cross Abstract: Detecting fraud in financial transactions typically relies on tabular models that demand heavy feature engineering to handle high-dimensional

v0.20.6

Local AiDGX agent

Ollama v0.20.6-rc0 is a pre-release update to the Ollama local model runner, published on April 10, 2026. Key changes include adding a Hermes agent integration guide to the docs, fixing missing par...

Validated Intent Compilation for Constrained Routing in LEO Mega-Constellations

Model ReleasesDGX agent

arXiv:2604.07264v1 Announce Type: cross Abstract: Operating LEO mega-constellations requires translating high-level operator intents ('reroute financial traffic away from polar links under 80 ms') int

VenusBench-Mobile: A Challenging and User-Centric Benchmark for Mobile GUI Agents with Capability Diagnostics

Model ReleasesDGX agent

arXiv:2604.06182v1 Announce Type: cross Abstract: Existing online benchmarks for mobile GUI agents remain largely app-centric and task-homogeneous, failing to reflect the diversity and instability of

VisCoder2: Building Multi-Language Visualization Coding Agents

Model ReleasesDGX agent

arXiv:2510.23642v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have recently enabled coding agents capable of generating, executing, and revising visualization code. However, e

Vision-Language Foundation Models for Comprehensive Automated Pavement Condition Assessment

ApplicationsDGX agent

arXiv:2604.08212v1 Announce Type: new Abstract: General-purpose vision-language models demonstrate strong performance in everyday domains but struggle with specialized technical fields requiring preci

Visual prompting reimagined: The power of the Activation Prompts

Model ReleasesDGX agent

arXiv:2604.06440v1 Announce Type: cross Abstract: Visual prompting (VP) has emerged as a popular method to repurpose pretrained vision models for adaptation to downstream tasks. Unlike conventional mo

we are doing this podcast specifically to highlight the innermost details of building agents in production may be a little niche, but its a …

ApplicationsDGX agent

we are doing this podcast specifically to highlight the innermost details of building agents in production may be a little niche, but its a niche i like 🤷‍♂️ @hwchase17 Great content, really appreciat

Weakly-Supervised Lung Nodule Segmentation via Training-Free Guidance of 3D Rectified Flow

ResearchDGX agent

arXiv:2604.08313v1 Announce Type: new Abstract: Dense annotations, such as segmentation masks, are expensive and time-consuming to obtain, especially for 3D medical images where expert voxel-wise labe

← Previous
1…163164165166
Next →