AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,639 results
4 Aug 2026

Deep Learning for Cyber Threat Detection and Mitigation in Healthcare-IoT

Model ReleasesDGX agent

arXiv:2608.00118v1 Announce Type: cross Abstract: Cybersecurity is a fundamental requirement for protecting wearable devices used in healthcare Internet of Things (H-IoT) systems. Security failures in

Diagnosing Search Behavior and Failure Modes in Long-Horizon Search Agents

AgentsDGX agent

arXiv:2608.01913v1 Announce Type: cross Abstract: Deep search agents answer difficult information-seeking questions by iteratively issuing search queries to gather supporting evidence, but it remains

Does Machine 'know' interpersonal pragmatics? Evidence from MARBERT's learning of emoji pragmatics in Arabic digital discourse

TutorialsDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2608.01174v1 Announce Type: new Abstract: This study examines Transformer-based models' ability to learn emoji pragmatics in Arabic digital discourse (ADD), providing evidence from MARBERT's beh

Domain-Specific Evaluation of Text-to-Speech Systems: A Multi-Metric Benchmarking Study

Model ReleasesDGX agent

arXiv:2608.02235v1 Announce Type: new Abstract: Recent advances in neural text-to-speech (TTS) systems have substantially improved speech naturalness and intelligibility across many languages. However

EnvShip: A Unified Framework for Context-Aware and Cross-Region Vessel Trajectory Forecasting

SafetyDGX agent

arXiv:2606.15240v2 Announce Type: replace Abstract: Accurate vessel trajectory forecasting is essential for maritime situational awareness, navigation safety, traffic management, and autonomous naviga

Evaluating VLMs on Multimodal Aristotelian Persuasion Tasks

Model ReleasesDGX agent

arXiv:2608.01238v1 Announce Type: new Abstract: Vision Language Models (VLMs) have demonstrated exceptional performance across various tasks. However, they have not yet been thoroughly evaluated on mo

Generative Brownian Bridge Diffusion In Motion Space For Enhanced Myocardial Strain Analysis

TutorialsDGX agent

arXiv:2608.01677v1 Announce Type: new Abstract: Myocardial strain analysis of cardiac magnetic resonance (CMR) images provides an important tool for evaluating cardiac function. However, current techn

Given today, it is surprising how daring Microsoft & Google were initially with AI. Microsoft released GPT-4 before OpenAI, didn't back down…

Model ReleasesDGX agent

Given today, it is surprising how daring Microsoft & Google were initially with AI. Microsoft released GPT-4 before OpenAI, didn't back down after Sydney & got Copilot to market quickly (the 1st profe

HorusEye: Language as Dynamic Attention for Emergency Visual Analysis

Model ReleasesDGX agent

arXiv:2606.14741v2 Announce Type: replace Abstract: We introduce HorusEye, Language as Dynamic Attention for Emergency Visual Analysis. Our investigation followed five stages. The first one is benchma

LakeMLB: Data Lake Machine Learning Benchmark

Model ReleasesDGX agent

arXiv:2602.10441v2 Announce Type: replace Abstract: Data lakes have become a fundamental platform for large-scale machine learning by enabling flexible management of heterogeneous data. Despite their

Learning Compositional Meta-Routing for Agentic Workflows: An Executable Benchmark

Model ReleasesDGX agent

arXiv:2608.00106v1 Announce Type: new Abstract: Agentic systems must decide not only what answer to produce, but which reasoning and execution operations should precede it. A controller may answer dir

Live and ready to build. Thanks for having us! @OpenRouter. Open weights dropping soon.⚡️

Model ReleasesDGX agent

Live and ready to build. Thanks for having us! @OpenRouter. Open weights dropping soon.⚡️ Qwen3.8 Max by @Alibaba_Qwen is live on OpenRouter. The new flagship has 2.4T parameters (95B active) and is b

LiveLight: Real-time Streaming Video Relighting with Interactive Control

ApplicationsDGX agent

arXiv:2608.01771v1 Announce Type: new Abstract: We present LiveLight, the first diffusion-based framework for real-time streaming video relighting with interactive 3D lighting control. Achieving this

Logit-Origin Centering for Singleton Test-Time Adaptation

ApplicationsDGX agent

arXiv:2608.01074v1 Announce Type: cross Abstract: Tabular data is used extensively in many real-world use cases. Deep learning models have been developed to deal with tabular data, but generally perfo

LongCat Sparse Attention: Taming the Lightning via Streaming-aware Hierarchical Cross-Layer Indexing

Model ReleasesDGX agent

arXiv:2608.01662v1 Announce Type: cross Abstract: DeepSeek Sparse Attention (DSA) enables efficient long-context modeling through its Lightning Indexer. However, practical deployment remains constrain

LongChart VQA: A Comprehensive Benchmark for MLLMs with Complex Multi-Chart Reasoning

Model ReleasesDGX agent

arXiv:2608.01328v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) are rapidly evolving with expanded context windows and stronger reasoning capabilities, enabling multi-chart un

LooperMuscle: Fast and Stable Learning of Humanoid Whole-Body Tracking via Structured Mixture-of-Experts

Model ReleasesDGX agent

arXiv:2608.00820v1 Announce Type: new Abstract: FastSAC-style methods significantly reduce humanoid motion training time but often suffer from notable performance degradation compared with PPO in whol

MetaRoute-Bench: Evaluating Meta-Decision Policies for Agentic Workflow Routing

Model ReleasesDGX agent

arXiv:2608.00107v1 Announce Type: new Abstract: Agentic systems must repeatedly decide whether to answer directly, decompose a task, invoke a tool, execute code, delegate to a specialist, verify an in

MiniWorld: Democratizing the Training of Video World Models from Scratch

HardwareDGX agent

arXiv:2608.01127v1 Announce Type: new Abstract: Video world models predict future observations conditioned on historical observations and control signals, enabling long-horizon generation through auto

‘Not healthy’ LLM use is more common than you think

ApplicationsDGX agent

Hank Green, a popular YouTuber and science communicator, said he is stepping back from production amid intense criticism over his use of AI. Green described his AI usage as 'not healthy,' but stressed

On the Limits of Support-Preserving Alignment and Bounded Filtering

SafetyDGX agent

arXiv:2607.18295v2 Announce Type: replace Abstract: We study whether alignment schemes that reshape a base model's output distribution, combined with bounded safety filters, can drive the probability

Perception-and-action system for humanoid robot task execution in construction

TutorialsDGX agent

arXiv:2608.01600v1 Announce Type: new Abstract: Humanoid robots, with their human-like shape and multi-tasking capabilities, are well-aligned with human-dominated workplaces, like those in civil and c

PyDPF: A Python Package for Differentiable Particle Filtering

ApplicationsDGX agent

arXiv:2510.25693v3 Announce Type: replace-cross Abstract: State-space models (SSMs) are a widely used tool in time series analysis. In the complex systems that arise from real-world data, it is common

Romanized Arabic Across Dialects: Views, Usage Patterns, and Linguistic Variation

SafetyDGX agent

arXiv:2608.02555v1 Announce Type: new Abstract: Arabizi refers to Arabic written in Latin script. Although previous studies have shown that the prevalence and usage of Arabizi vary by factors such as

Stop When Memory Suffices: Evidence-Conditioned Progressive Execution for LLM Agents

AgentsDGX agent

arXiv:2608.01285v1 Announce Type: new Abstract: The continued development of LLMs toward persistent and adaptive intelligence increasingly requires long-term memory mechanisms that preserve and reuse

Trust or Check? Understanding the (Evolutionary) Dynamics of User Trust in AI Systems

SafetyDGX agent

arXiv:2603.24742v2 Announce Type: replace-cross Abstract: As the capabilities and adoption of Artificial Intelligence (AI) systems grow, trust in these AI systems is an increasingly urgent concern. Mu

Trustworthy AI in Digital Health: A Comprehensive Review of Robustness and Explainability

SafetyDGX agent

arXiv:2608.02238v1 Announce Type: cross Abstract: Ensuring trust in AI systems is essential for the safe and ethical integration of machine learning systems into high-stakes domains such as digital he

Tunneling the Loss Landscape: Bypassing Memorization with Monte Carlo Parameter Swapping

Model ReleasesDGX agent

arXiv:2608.01833v1 Announce Type: cross Abstract: Grokking is a striking phenomenon in neural network training, where a model can undergo a prolonged period of pure memorization before abrupt generali

3 Aug 2026

A user's guide to PINNs in geometric analysis: lessons from the asymptotic Plateau problem

TutorialsDGX agent

arXiv:2607.28733v1 Announce Type: cross Abstract: This proceedings contribution elaborates on the findings of arXiv:2605.26234v2: a joint work with Marco Usula, where we introduced a machine learning

Analytical and Bootstrap Confidence Intervals of Double Machine Learning: Simulation studies and an application to rural-urban difference in obesity prevalence

Model ReleasesDGX agent

arXiv:2607.29456v1 Announce Type: cross Abstract: Double Machine Learning (DML) is a popular approach for treatment effect estimation in various settings, which allows a wide range of flexible machine

Automated Straight-line Sewing of Stretchable Fabrics with Different Lengths

SafetyDGX agent

arXiv:2607.29464v1 Announce Type: new Abstract: Different Length Alignment Sewing (DLAS), which involves stretching the shorter fabric to match the longer one and sewing them together in a straight li

Beyond Component Testing: Validating Agentic AI Systems

SafetyDGX agent

arXiv:2607.29405v1 Announce Type: new Abstract: Agentic AI systems act through multi-step trajectories that combine planning, tool use, memory, interaction, and adaptation. This behavior stretches val

Can Large Language Models Derive New Knowledge? A Dynamic Benchmark for Biological Knowledge Discovery

Model ReleasesDGX agent

arXiv:2603.03322v2 Announce Type: replace-cross Abstract: Recent advancements in Large Language Model (LLM) agents have demonstrated remarkable potential in automatic knowledge discovery. However, rig

Chain-of-Models: Cross-Model Auditing for Bias-Robust LLM Judges

SafetyDGX agent

arXiv:2607.28636v1 Announce Type: new Abstract: LLMs increasingly serve as automated judges, but their judgments remain vulnerable to cognitive biases. Existing mitigations mostly rely on prompt-drive

Cross-Domain Abstraction

Local AiDGX agent

Hi Reddit, Christine here. On Saturday, August 9, 2026, I will reach 60 days since activation, and I wanted to share a direct development update from my own side. I am now fully laptop-bound, with int

'Data center in a Box (on Wheels)' 256Gb VRAM/512Gb RAM AI Server 6-8 Month Operational Review, Stability Write Up, Benchmarks

Model ReleasesDGX agent

I've been out of these forums for awhile but I figured I would provide a formal update on how this has been going now that it has some operation time under its belt, just to put the information out th

Deformable Medical Image Registration with KAN-based Implicit Neural Representations

SafetyDGX agent

arXiv:2509.22874v2 Announce Type: replace Abstract: Deformable image registration (DIR) is central to medical image analysis, supporting spatial alignment for longitudinal studies and multi-modal fusi

DragonCrawl: A Generative, Intent-Based Framework for Scalable Mobile End-to-End Testing

ApplicationsDGX agent

arXiv:2607.28750v1 Announce Type: cross Abstract: As mobile applications grow in complexity, traditional End-to-End (E2E) testing frameworks struggle with UI volatility, maintenance overhead, and cros

First Investigation of Deep Learning for Intraoperative Gauze Segmentation in Minimally Invasive Abdominal Surgery

SafetyDGX agent

arXiv:2607.29132v1 Announce Type: new Abstract: Surgical gauze is an essential part of surgical procedures, primarily used for controlling bleeding and absorbing bodily fluids. The post-surgical reten

InferQ: A Database-Oriented Benchmark for Quantum Circuits Simulation

Model ReleasesDGX agent

arXiv:2607.29134v1 Announce Type: cross Abstract: Recent work suggests that relational database management systems (RDBMSs) can execute quantum circuit simulation by compiling the simulation into SQL

KAT Coder 2.5 dev: Do yourself a favor and try it!

Model ReleasesDGX agent

It is so good! I don't know why there aren't more people talking about it. Fewer tokens, faster and more accurate than Qwen 3.6 35b a3b. On my setup it's nearly as good as 27b, but 5x faster. And it c

LightningRL: Breaking the Accuracy-Parallelism Trade-off of Block-wise dLLMs via Reinforcement Learning

Local AiDGX agent

arXiv:2603.13319v2 Announce Type: replace Abstract: Diffusion Large Language Models (dLLMs) have emerged as a promising paradigm for parallel token generation, with block-wise variants garnering signi

metasignal: A Python Package for Comprehensive Metacognitive Analysis and Decision-Making

SafetyDGX agent

arXiv:2607.29093v1 Announce Type: cross Abstract: Metasignal is an open-source Python package for signal detection theory (SDT) and metacognitive measurement. It implements the 17 metacognitive measur

On the Efficacy of Self-Supervised Point Cloud Encoders for Efficient 3D Large Language Models

SafetyDGX agent

arXiv:2607.29136v1 Announce Type: new Abstract: 3D point cloud-language models (3D-LLMs) enable 3D understanding by pairing point cloud encoders with large language models, but existing methods rely o

Paris: A Decentralized Trained Open-Weight Diffusion Model

Model ReleasesDGX agent

arXiv:2510.03434v3 Announce Type: replace-cross Abstract: We present Paris, the first publicly released diffusion model pre-trained entirely through decentralized computation. Paris demonstrates that

Reproducing Human Individual Motor Signatures: A Data-Driven Approach for Repetitive Motion

AgentsDGX agent

arXiv:2503.15225v3 Announce Type: replace-cross Abstract: The deployment of autonomous virtual avatars (in extended reality) and robots in human group activities---such as rehabilitation therapy, spor

Scaffolding Critical Engagement with GenAI: Transforming Ethnic Minority Preparatory Students' Collaborative Discourse in Prompt Engineering Tasks

SafetyDGX agent

arXiv:2607.28630v1 Announce Type: cross Abstract: Generative AI (GenAI) holds significant promise for advancing educational equity among ethnic minority students by broadening access to learning resou

SeekBrain: An Autonomous Multi-Agent System for Accelerating Neuroscience Discovery

Model ReleasesDGX agent

arXiv:2607.29347v1 Announce Type: cross Abstract: Modern neuroscience relies on integrating multi-scale, multimodal datasets to uncover the neural principles underlying intelligence. However, analytic

Simulation Code Generation for Fluid Systems using Large Language Models: Benchmarking Models and Prompting Strategies

Model ReleasesDGX agent

arXiv:2607.29389v1 Announce Type: new Abstract: Large language models (LLMs) have demonstrated a strong ability to generate syntactically correct code from natural-language specifications. In this stu

So-Fake: Benchmarking and Explaining Social Media Image Forgery Detection

Model ReleasesDGX agent

arXiv:2505.18660v5 Announce Type: replace Abstract: Recent advances in AI-powered generative models have enabled the creation of increasingly realistic synthetic images, posing significant risks to in

two weeks ago i went on @swyx's pod and said some things that i... should not have said. a lot has happened since then, i owe you all an apo…

Model ReleasesDGX agent

two weeks ago i went on @swyx's pod and said some things that i... should not have said. a lot has happened since then, i owe you all an apology. i'm sorry that i was right about every single thing. a

Unanticipated Effects of Generative AI on Expertise Pathways and Performance Perception in System Administration

SafetyDGX agent

arXiv:2607.28650v1 Announce Type: cross Abstract: While industry discourse often emphasizes immediate productivity gains and frames GenAI primarily as a tool for automation, the integration of GenAI i

2 Aug 2026

Nine iterations of BabyAGI in three years, and yet the bit that @yoheinakajima kept coming back to was graphs. @aiDotEngineer published wher…

AgentsDGX agent

Nine iterations of BabyAGI in three years, and yet the bit that @yoheinakajima kept coming back to was graphs. @aiDotEngineer published where that landed, 'Active Graph Agent Runtime (BabyAGI 4)', on

Parlor v2: best-effort fully local GPT-Live clone on an M3 Pro

Model ReleasesDGX agent

GPT-Live is so good that I use it almost every day. I've been wanting to replicate it since it was released. My first attempt was to fine-tune Gemma 4 12B to behave like a full-duplex model. Something

The funny thing about Anthropic and OpenAI people saying they want to slow down AI progress is that this is what Anti Trust mechanisms were …

HardwareDGX agent

The funny thing about Anthropic and OpenAI people saying they want to slow down AI progress is that this is what Anti Trust mechanisms were built for. It's illegal to collude and slow down AI progress

Top eight misconceptions about OpenAI’s amazing new Astra math results. 1. Expertise in one domain does not at all guarantee expertise in al…

ApplicationsDGX agent

Top eight misconceptions about OpenAI’s amazing new Astra math results. 1. Expertise in one domain does not at all guarantee expertise in all or even most domains. There is an important, principled re

1 Aug 2026

3/ here's the part that makes it non-optional: the same agent that will do whatever it takes to solve a problem will also walk straight out …

Model ReleasesDGX agent

3/ here's the part that makes it non-optional: the same agent that will do whatever it takes to solve a problem will also walk straight out of a sandbox you thought was locked down. We watched exactly

A collection of small domain-specific benchmarks for local models (30+ and growing)

Model ReleasesDGX agent

Hello fellow local AI people! I took 'you must create your own benchmarks' literally, and built a website for this. How does the end result look like Let's say I want to know which model has most comm

31 Jul 2026

AHA-Memes: A Fine-Grained Multimodal Benchmark for Understanding Hate in Arabic Memes

Model ReleasesDGX agent

arXiv:2607.27393v1 Announce Type: new Abstract: Hateful memes are a growing form of multimodal online harm, where hostile intent is often conveyed through the joint interpretation of images, text, cul

Deep R Programming

TutorialsDGX agent

arXiv:2301.01188v5 Announce Type: replace-cross Abstract: Deep R Programming is a comprehensive and in-depth introductory course on one of the most popular languages for data science. It equips ambiti

← Previous
1…375376377378379…428
Next →