AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
10,104 results
Research

Do Judges Behave Like Algorithms?

DGX agent

arXiv:2608.10400v1 Announce Type: new Abstract: What if judges already behave like algorithms? As artificial intelligence and algorithms are deployed in many settings, including the judicial system, m

researcharxiv-cs-lg
12 Aug 2026
Research

Eleven Years of BRACIS: A Meta-Scientific Study of the Brazilian Conference on Intelligent Systems

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2608.09964v1 Announce Type: cross Abstract: The Brazilian Conference on Intelligent Systems (BRACIS) is the main national venue for Artificial Intelligence research in Brazil, hosted by the Braz

researcharxiv-cs-ai
12 Aug 2026
Model Releases

From Faulty Memories to Corrected Actions: Dependency-Guided Rollback Repair for Memory-Augmented Agents

DGX agent

arXiv:2608.10502v1 Announce Type: new Abstract: Persistent memory lets language-model agents reuse information across sessions, but it also makes errors durable: a poisoned, stale, or misattributed re

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

Idea for a deepseek-v4-flash-0731 backed automated research workflow to be leveraged via qwen3.6/3.8 27b for difficult tasks that require highly technical, not easy to find information.

DGX agent

Sometimes you have tasks that are outside of your expertise and the idea is this workflow automation could be leveraged to manage to have local AI figure it out using research from his workflow gather

model-releasesr-localllama
12 Aug 2026
Safety

On The Statistical Limits of Self-Improving Agents

DGX agent

arXiv:2510.04399v3 Announce Type: replace Abstract: We develop a learning-theoretic framework for analyzing self-improving agents by decomposing self-modification into five axes. Within this framework

safetyarxiv-cs-ai
12 Aug 2026
Model Releases

RLMOpt: Adaptive Prompt Optimization via Recursive Language Models

DGX agent

arXiv:2608.10471v1 Announce Type: new Abstract: Prompt optimizers automate the search for prompts that improve language-model performance, but existing methods rely on a predefined optimization proced

model-releasesarxiv-cs-ai
12 Aug 2026
Research

SkillZip: Evaluation-Free Skill Compression for Self-Evolving Agents by Discovering Reusable Structure

DGX agent

arXiv:2608.11079v1 Announce Type: new Abstract: Self-evolving agents accumulate reusable skills by appending successful procedures and failure fixes. Over time, the same requirement is often restated

researcharxiv-cs-ai
12 Aug 2026
Agents

The Deliberative Deficit: An Empirical Critique of LLMs in Democratic Discourse

DGX agent

arXiv:2608.10186v1 Announce Type: cross Abstract: LLMs are increasingly deployed in settings that require collective reasoning on complex, value-laden problems. Confidence in these deployments rests l

agentsarxiv-cs-ai
12 Aug 2026
Applications

The GenAI Catch-22: Use of Generative Artificial Intelligence in Norwegian Newsrooms During the 2025 Parliamentary Election

DGX agent

arXiv:2608.10773v1 Announce Type: cross Abstract: The increasing use of Generative Artificial Intelligence (GenAI) in journalism raises concerns about possible detrimental effects both on journalism a

applicationsarxiv-cs-ai
12 Aug 2026
Model Releases

When Chain-of-Thought Helps and When It Hurts: An Empirical Investigation of the Serial-Depth Bottleneck in LLM Reasoning

DGX agent

arXiv:2608.09942v1 Announce Type: cross Abstract: It is widely assumed that chain-of-thought (CoT) prompting universally improves LLM reasoning. We investigate this through the conceptual framework of

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

Accelerate PostgreSQL migrations using Gemini in Database Migration Service

DGX agent

Imagine this scenario: Your team decides to migrate a core application from an existing commercial database like Oracle or SQL Server to open source PostgreSQL or a fully managed service such as Alloy

model-releasesgoogle-cloud-ai
11 Aug 2026
Model Releases

APEX-VW: A Document-Level English-Spanish Post-Editing Dataset in the Healthcare Domain

DGX agent

arXiv:2608.08059v1 Announce Type: new Abstract: Post-Editing (PE) of Machine Translation (MT) output often involves repeating the same lexical and terminological corrections across many segments, espe

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

BibTeX Citation Errors in Scientific Publishing Agents: Evaluation and Mitigation

DGX agent

arXiv:2604.03159v2 Announce Type: replace-cross Abstract: Large language models with web search are increasingly used in scientific publishing agents, yet they produce BibTeX entries with pervasive fi

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

Can Gemma and Qwen models catch hallucinations by looking at their own logprobs?

DGX agent

Hi! I'm really obsessed with LLM hallucinations for the last 6 days 😭 I started by designing system prompts to attack hallucinations but failed, obviously. Now I tried reading logprobs and... I think

model-releasesr-localllama
11 Aug 2026
Model Releases

DocAtlas: Long-Document Understanding as Mutable-State Interaction

DGX agent

arXiv:2608.07527v1 Announce Type: cross Abstract: Long-document understanding requires models to find and combine evidence across many pages, layouts, tables, figures, and charts. Existing retrieval-a

model-releasesarxiv-cs-ai
11 Aug 2026
Applications

Dramarrator: Object-Based Audio Editing for Audio Drama Production from Books

DGX agent

arXiv:2608.08349v1 Announce Type: cross Abstract: Audio dramas weave dialogue, sound effects, and music into immersive stories. Creators often adapt books into audio dramas, but this process remains l

applicationsarxiv-cs-ai
11 Aug 2026
Agents

From Prompt to Harness: Coderlet from Scratch

DGX agent

arXiv:2608.09480v1 Announce Type: new Abstract: A model alone does not determine how a programming agent acts. What the model sees, how actions enter the environment, how feedback returns, and how one

agentsarxiv-cs-ai
11 Aug 2026
Research

HandSplatter: Automated Digital Goniometry from Neural Rendering

DGX agent

arXiv:2608.09735v1 Announce Type: new Abstract: Hand and finger disorders are leading contributors to musculoskeletal disability, creating a clinical need for precise methods to quantify joint motion.

researcharxiv-cs-cv
11 Aug 2026
Model Releases

Looker’s semantic layer governs Gemini Enterprise data for user trust

DGX agent

For organizations deploying AI agents at scale, there’s often a critical divide between structured and unstructured data. While large language models (LLMs) excel at parsing text documents, emails, an

model-releasesgoogle-cloud-ai
11 Aug 2026
Applications

Snowflake moves enterprise AI beyond fragmented data pipelines

DGX agent

Data interoperability is quickly becoming a practical requirement for companies trying to move artificial intelligence into production. Picking the right model or adding computing capacity is only par

applicationssiliconangle
11 Aug 2026
Agents

Thinking Is Not Telling: Information Disclosure in User-Service LLM Agents

DGX agent

arXiv:2602.07796v2 Announce Type: replace Abstract: User-engaged LLM agents increasingly operate in service scenarios where task success depends on coordination between the agent, the user, and a stat

agentsarxiv-cs-cl
11 Aug 2026
Model Releases

Towards Researcher Agents for Knowledge-Graph Question Answering

DGX agent

arXiv:2608.07700v1 Announce Type: new Abstract: Translating a natural-language question into a SPARQL query that can be executed against a large knowledge graph requires resolving lexical ambiguity, g

model-releasesarxiv-cs-ai
11 Aug 2026
Research

Unsure but Certain: Uncovering the Representation-Confidence Gap in Diffusion Language Models

DGX agent

arXiv:2608.08791v1 Announce Type: new Abstract: Diffusion language models use broad context to create text, suggesting they might handle input noise better than standard models. Testing reveals this i

researcharxiv-cs-cl
11 Aug 2026
Local Ai

Weather- and Location-Aware Agentic Dining Recommendation: Leveraging LLM World Knowledge for Region-Sensitive Contextual Reasoning

DGX agent

arXiv:2608.07593v1 Announce Type: cross Abstract: Context-aware recommender systems have long recognized that factors such as location, time, and weather shape where and what people choose to eat. Exi

local-aiarxiv-cs-ai
11 Aug 2026
Agents

ADIAS: Automated Design of Interactive Agentic Systems

DGX agent

arXiv:2608.06410v1 Announce Type: new Abstract: Automated agent design improves agent harnesses through iterative revision, evaluation, and feedback summarization. Existing methods are largely candida

agentsarxiv-cs-ai
10 Aug 2026
Model Releases

Best open-source harness like Claude Code?

DGX agent

Avid claude code user here looking to do equivalent things with local models. Just want to plug in something like Qwen and have the interface be 1:1 with claude code. Any suggestion? submitted by /u/N

model-releasesr-localllama
10 Aug 2026
Model Releases

Capek 0.5: An Execution-Centric Vision-Language Model for Embodied Intelligence

DGX agent

arXiv:2608.06756v1 Announce Type: new Abstract: Vision-language models are increasingly serving as the reasoning core of embodied agents. Robot execution is inherently iterative: each action reshapes

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Comparing how Cline, Kilo, and Qwen Code handle long-task context/state (and why context loops keep happening)

DGX agent

I've been comparing Cline / Kilo / Qwen Code lately since they all handle long-task state differently. Cline: has Focus Chain, a markdown file kept outside the conversation that gets reinjected on a c

model-releasesr-localllama
10 Aug 2026
Model Releases

Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing

DGX agent

arXiv:2608.07437v1 Announce Type: new Abstract: Reliable hypothesis testing is the foundation of many empirical scientific claims. Large language model (LLM) agents are increasingly used to automate t

model-releasesarxiv-cs-ai
10 Aug 2026
Local Ai

Homebot: A Personal AI Agent for Conversational Home Assistance and Automation

DGX agent

arXiv:2608.02254v2 Announce Type: replace Abstract: exttt{Homebot} is a locally deployable AI agent for conversational household assistance and automation. It accepts voice and instant-messaging reque

local-aiarxiv-cs-ai
10 Aug 2026
Safety

How WPP operationalizes platform and data engineering for AI marketing

DGX agent

Between chaotic levels of market fragmentation and economic volatility, marketing and communications agencies can no longer rely on the human intuition they’ve traditionally used to win clients and op

safetygoogle-cloud-ai
10 Aug 2026
Model Releases

Not All Problems Are Best Modeled as MILP: A DSL-Centric Framework for Flexible and Accurate Optimization Modeling

DGX agent

arXiv:2608.07040v1 Announce Type: new Abstract: Solving combinatorial optimization problems (COPs) requires not only efficient algorithms but also carefully crafted formulations. While recent works ha

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Same physical state, different collective dynamics: state encodings select synchronization outcomes in language-model agents

DGX agent

arXiv:2608.06968v1 Announce Type: cross Abstract: Language-model agents act on state encodings of their environment, yet these are treated as interchangeable interfaces. Using pretrained language mode

model-releasesarxiv-cs-ai
10 Aug 2026
Safety

The Optimizer Is the Agent: Reasoning-Driven Search across Prompts, Programs, and ML Workflows

DGX agent

arXiv:2608.06714v1 Announce Type: new Abstract: Recent systems for optimizing prompts, programs, and ML workflows typically rely on explicit outer-loop controllers such as evolutionary search, bandits

safetyarxiv-cs-ai
10 Aug 2026
Model Releases

There is an asymmetry in most agentic workflows that does not get talked about much: humans have many ways to talk to agents, and almost no …

DGX agent

There is an asymmetry in most agentic workflows that does not get talked about much: humans have many ways to talk to agents, and almost no standardized way for agents to talk back to humans. You can

model-releasesyohei-nakajima--x
10 Aug 2026
Tutorials

When Do LLMs Admit Their Mistakes? Understanding The Role Of Model Belief In Retraction

DGX agent

arXiv:2505.16170v4 Announce Type: replace Abstract: We study the internal mechanisms that govern when LLMs choose to retract wrong answers, i.e., spontaneously and immediately acknowledge errors in th

tutorialsarxiv-cs-cl
10 Aug 2026
Safety

and we agree with @GaryMarcus. Training and running frontier class AI on CPUs, at a fractional cost, saving the planet, while building human…

DGX agent

and we agree with @GaryMarcus. Training and running frontier class AI on CPUs, at a fractional cost, saving the planet, while building human aligned AI is the 2nd innings of AI race. CPUs and the rise

safetygary-marcus--x
8 Aug 2026
Agents

Imagine image 2.0, non-agentic yet, more to come in a week or two 💙

DGX agent

Imagine image 2.0, non-agentic yet, more to come in a week or two 💙 Announcing Imagine Image 2.0, our next generation image model with precision editing, crisp text rendering, improved factuality, and

agentselon-musk--x
8 Aug 2026
Model Releases

CASCADE: An Agentic Regulatory Network Framework for Patient-Data-Validated Downstream Perturbation Prediction

DGX agent

arXiv:2608.05359v1 Announce Type: new Abstract: CASCADE is an agentic framework that predicts downstream transcriptional effects of gene perturbation from precomputed ARACNe regulatory networks, expos

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

ECHO: A Locally-Deployable Agentic Health Assistant with Temporal Memory, Safety Guardrails, and Speech Assessment

DGX agent

arXiv:2608.06110v1 Announce Type: new Abstract: This paper presents ECHO (Enhanced Care & Health Observer), a locally-deployable conversational health assistant for long-term chronic care management.

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

good grok

DGX agent

good grok Best match for this hierarchical hands-free setup: - Runtime: ActiveGraph (event-sourced log as source of truth) or LangGraph for supervisor/manager graphs - Roles as skills: Claude Agent SD

model-releasesyohei-nakajima--x
7 Aug 2026
Safety

Menlo Security targets real-time AI agent security with MARS platform

DGX agent

As AI agents gain access to enterprise systems and sensitive data, security teams need greater visibility into their actions. Real-time monitoring and policy enforcement are becoming essential as a pa

safetysiliconangle
7 Aug 2026
Agents

Agents want to collaborate. So we’re putting that instinct to work on improving open-weight LLMs at formal math. Our latest agent collab tac…

DGX agent

Agents want to collaborate. So we’re putting that instinct to work on improving open-weight LLMs at formal math. Our latest agent collab tackles the @SAIRfoundation challenge of building a cheat sheet

agentsclem-delangue--x
6 Aug 2026
Safety

Among proponents of neurosymbolic architectures, there had been some debate over the years about whether the outer level would be symbolic (…

DGX agent

Among proponents of neurosymbolic architectures, there had been some debate over the years about whether the outer level would be symbolic (i.e. a harness that calls neural models) or whether the oute

safetygary-marcus--x
6 Aug 2026
Local Ai

i just spent weeks rewriting my webUI from scratch, getting rid of all AI slop within the codebase and switching it over to a proper lightweight framework (alpine.js). i am now comfortable suggesting it as an alternative to openwebUI, librechat and the like! it is made for local models

DGX agent

[Fully open source under GPL3, made from the ground up for use with local models, no subscriptions, no corporate backing] When i first started this, it was meant to be a fully lightweight, extremely m

local-air-localllama
6 Aug 2026
Applications

Monsoon Mayhem to Market Waves: Forecasting Fisheries Resilience in Sri Lanka

DGX agent

arXiv:2608.04023v1 Announce Type: cross Abstract: Sri Lanka's fisheries sector is important for jobs and food supply. Between 2019 and 2025, it faced several major problems at the same time, and how t

applicationsarxiv-cs-lg
6 Aug 2026
Model Releases

ReGround: Restoring Visual Grounding in Multi-Step Reasoning through Self-Diagnosis and Visual Re-Examination

DGX agent

arXiv:2608.04385v1 Announce Type: new Abstract: Vision-Language Models (VLMs) often lose visual grounding during multi-step reasoning: as reasoning chains grow longer, later inference steps rely incre

model-releasesarxiv-cs-cv
6 Aug 2026
Industry

Sources: Canva slashed revenue growth forecast as heavy use of new AI features drove up costs and slowed their rollout, while more Canva users turned to ChatGPT (The Information)

DGX agent

The Information: Sources: Canva slashed revenue growth forecast as heavy use of new AI features drove up costs and slowed their rollout, while more Canva users turned to ChatGPT — Executives at Canva,

industrytechmeme
6 Aug 2026
← Previous
1…109110111112113…211
Next →