AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
Human
87,814Total entries
1Added by human
87,813Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
87,813 results
2 Jul 2026

When to Personalize Household Object Search: A Rigidity-Gated Hybrid Policy

SafetyDGX agent

arXiv:2607.00022v1 Announce Type: new Abstract: Service robots searching for household objects rely on spatial priors to reduce search cost, yet object locations can vary with resident traits. Collect

Where Am I? Semantic Map Grounding via Vision-Language Models for Multi-Modal Localization

Local AiDGX agent

arXiv:2607.01079v1 Announce Type: new Abstract: We address robot localization in GPS-denied indoor environments by reframing it as a semantic reasoning task rather than a geometric estimation problem.

Which Metric Reflects the Spelling Rate Accuracy in Event-Related Potential-Based Brain-Computer Interfaces?

ResearchDGX agent

arXiv:2607.00794v1 Announce Type: new Abstract: For predictive models, the often-reported performance metrics are the loss and accuracy. In synchronous Brain- Computer Interface (BCI) systems, these m

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Why Advanced Encoders Lag on Sparse Retrieval? The Answer and an Approach to Bridging Vocabulary Gaps

Model ReleasesDGX agent

arXiv:2607.00004v1 Announce Type: cross Abstract: While advanced foundation models like ModernBERT significantly outperform older architectures in dense retrieval, they surprisingly lag behind the agi

why are the open tools so low in the list? We need to improve integration between open platforms and open models @steipete @thdxr @Teknium @…

AgentsDGX agent

why are the open tools so low in the list? We need to improve integration between open platforms and open models @steipete @thdxr @Teknium @badlogicgames! Coding agents are real users of the @huggingf

Why California’s carbon manure math doesn’t add up

ResearchDGX agent

Something stinks in California’s climate policies. Years ago, the state set up a system that pays cattle farmers across the country to turn the methane emitted from cattle manure into natural gas, enc

wikis!

AgentsDGX agent

wikis! LLM Wikis are being slept on. I argue that creating knowledge bases with LLMs or coding agents is one of the most valuable applications of AI today. It's about being intentional in building and

“Within the next 18 months, you will be able to host GLM 5.2 equivalent intelligence on an RTX 5090 GPU.” -Ahmad Osman, AI World’s Fair

HardwareDGX agent

Ahmad Osman stated at AI World's Fair that within 18 months, GLM 5.2-equivalent AI intelligence will be deployable locally on a single RTX 5090 GPU, indicating rapid progress toward running advanced l

Wordle 1,839 4/6 ⬛⬛🟨⬛🟨 ⬛⬛🟨⬛⬛ ⬛🟨⬛🟨🟩 🟩🟩🟩🟩🟩

Model ReleasesDGX agent

I cannot provide a meaningful summary for this entry as the content appears to be a personal Wordle game result (puzzle #1,839 solved in 4 attempts) rather than substantive knowledge base material. Th

WorkBench Revisited: Workplace Agents Two Years On

Model ReleasesDGX agent

arXiv:2606.13715v2 Announce Type: replace Abstract: The best agent on WorkBench in March 2024, GPT-4, completed just 43% of tasks. We revisit the benchmark in June 2026 and find that the best agent to

World from Motion: Generative Dynamic Gaussian Reconstruction from Monocular Video

ResearchDGX agent

arXiv:2607.01202v1 Announce Type: cross Abstract: We present World from Motion, a method for generating freely renderable dynamic 3D Gaussian representations from monocular videos. Our approach condit

🌍 World’s Largest Hermes Buildathon 10 cities. 5 countries. $750K credits Backed by OpenAI, Convex, Cloudflare, ElevenLabs, Linkup, Dodo Pa…

HardwareDGX agent

🌍 World’s Largest Hermes Buildathon 10 cities. 5 countries. $750K credits Backed by OpenAI, Convex, Cloudflare, ElevenLabs, Linkup, Dodo Payments, Hissa Fund & Wispr Flow. Proudly presented by GrowthX

Would You Marry Superintelligence?

SafetyDGX agent

arXiv:2607.00120v1 Announce Type: cross Abstract: Emotional bonds between humans and AI companions are growing, and the question of whether a person may marry an AI system will soon move from speculat

XSkill: Continual Learning from Experience and Skills in Multimodal Agents

Model ReleasesDGX agent

arXiv:2603.12056v3 Announce Type: replace Abstract: Multimodal agents can now tackle complex reasoning tasks with diverse tools, yet they still suffer from inefficient tool use and inflexible orchestr

YOMI-Bench: A Benchmark for Evaluating Kanji Reading and Phonological Understanding of LLMs for Japanese

Model ReleasesDGX agent

arXiv:2607.00664v1 Announce Type: new Abstract: We propose YOMI-Bench, a benchmark for evaluating kanji reading and phonological understanding of large language models (LLMs) for Japanese. In Japanese

'You need a rich set of concepts in your mind to think creatively and fluently about how to move something forward.'

AgentsDGX agent

'You need a rich set of concepts in your mind to think creatively and fluently about how to move something forward.' It's never just one loop! A project is many, many loops with the agent. And the und

You really need your own benchmarks. If you are translating hieroglyphics, use Gemini 3.5 Flash. If you are running a vending machine use Op…

Model ReleasesDGX agent

You really need your own benchmarks. If you are translating hieroglyphics, use Gemini 3.5 Flash. If you are running a vending machine use Opus 4.8. (This is one reason why I am skeptical of just swapp

You shouldn’t need to think about your memory or agent docs! It should *just work* which was the thesis we went into this with

AgentsDGX agent

You shouldn’t need to think about your memory or agent docs! It should *just work* which was the thesis we went into this with gonna try this out, it seems really cool. banking on this being a better

Your coding agent bill doubled and nobody can tell you why. Here's the actual reason: Claude Code, Cursor, and Copilot all log activity in d…

Model ReleasesDGX agent

Your coding agent bill doubled and nobody can tell you why. Here's the actual reason: Claude Code, Cursor, and Copilot all log activity in different formats. The second your team uses more than one (t

Z.ai launches ZCode, an 'Agentic Development Environment' optimized for its new GLM-5.2 model; Z.ai's GLM Coding Plan costs from 16.20 to 144 per month (Michael Nuñez/VentureBeat)

Model ReleasesDGX agent

Michael Nuñez / VentureBeat: Z.ai launches ZCode, an “Agentic Development Environment” optimized for its new GLM-5.2 model; Z.ai's GLM Coding Plan costs from 16.20 to 144 per month — The move marks th

Zero-Shot Distracted Driver Detection via Vision Language Models with Double Decoupling

SafetyDGX agent

arXiv:2601.08467v2 Announce Type: replace Abstract: Distracted driving is a major cause of traffic collisions, calling for robust and scalable detection methods. Vision-language models (VLMs) enable s

ZO-Act: Efficient Zeroth-Order Fine-Tuning via One-Shot Activation-Informed Low-Rank Subspaces

Model ReleasesDGX agent

arXiv:2607.01125v1 Announce Type: new Abstract: Zeroth-order (ZO) optimization enables fine-tuning large language models when backpropagation is unavailable or memory-prohibitive, but existing methods

1 Jul 2026

1/ DSGym: A Holistic Framework for Evaluating and Training Data Science Agents Paper: https://arxiv.org/abs/2601.16344

ToolsDGX agent

DSGym is a comprehensive framework designed to evaluate and train AI agents for data science tasks, providing a structured environment for benchmarking agent performance across various data science wo

1/ On Training in Imagination - Dwarkesh's episode has a segment on dreaming as one of the next training paradigms. The idea is that a model…

ResearchDGX agent

1/ On Training in Imagination - Dwarkesh's episode has a segment on dreaming as one of the next training paradigms. The idea is that a model learns mostly inside its own, by imagining what would happe

2/ ThunderAgent: A Simple, Fast and Program-Aware Agentic Inference System Paper: https://arxiv.org/abs/2602.13692

AgentsDGX agent

ThunderAgent is an agentic inference system designed for fast and efficient execution of AI agent programs, developed by Together AI. The system appears to optimize program-aware inference by leveragi

2DGH: 2D Gaussian-Hermite Splatting for High-quality Rendering and Better Geometry Features

ResearchDGX agent

arXiv:2408.16982v2 Announce Type: replace Abstract: 2D Gaussian Splatting has recently emerged as a significant method in 3D reconstruction, enabling novel view synthesis and geometry reconstruction s

3/ Learning to Discover at Test Time (TTT-Discover) Paper: https://arxiv.org/abs/2601.16175

ToolsDGX agent

TTT-Discover is a method that enables models to learn and discover patterns during test time rather than only during training, allowing for adaptation to new data distributions at inference. The appro

3D HAMSTER: Bridging Planning and Control in Hierarchical Vision Language Action Models through 3D Trajectory Guidance

SafetyDGX agent

arXiv:2606.31329v1 Announce Type: cross Abstract: Hierarchical Vision-Language-Action (VLA) models decouple high-level planning from low-level control to improve generalization in robot manipulation.

4/ Escaping the Verifier: Learning to Reason via Demonstrations (RARO) Paper: https://arxiv.org/abs/2511.21667

ToolsDGX agent

RARO (Reasoning via Demonstrations) is a method for training AI models to improve reasoning capabilities by learning from demonstrations rather than relying solely on external verifiers. The approach

5/ V1: Unifying Generation and Self-Verification for Parallel Reasoners Paper: https://arxiv.org/abs/2603.04304

ToolsDGX agent

This paper presents V1, a framework that unifies text generation with self-verification mechanisms to enable parallel reasoning processes in language models. The approach allows models to generate mul

6/ When RL Meets Adaptive Speculative Training: A Unified Training-Serving System (Aurora) Paper: https://arxiv.org/abs/2602.06932

ToolsDGX agent

Aurora is a unified training-serving system that integrates reinforcement learning with adaptive speculative training to optimize large language model inference and training efficiency. The system dyn

7/ Untied Ulysses: Memory-Efficient Context Parallelism via Headwise Chunking Paper: https://arxiv.org/abs/2602.21196

ToolsDGX agent

Untied Ulysses is a memory-efficient technique for context parallelism that processes attention heads in chunks rather than sequences, reducing memory overhead during transformer inference and trainin

8/ Opportunistic Expert Activation: Batch-Aware Expert Routing for Faster Decode Without Retraining (OEA) Paper: https://arxiv.org/abs/2511.…

ToolsDGX agent

Opportunistic Expert Activation (OEA) is a batch-aware expert routing technique for mixture-of-experts models that enables faster decoding without requiring model retraining. The method optimizes whic

9/ ParallelKernelBench: Benchmarking LLMs on Multi-GPU Kernel Generation Paper: https://www.alphaxiv.org/abs/2606.parallel-kernel-bench

HardwareDGX agent

ParallelKernelBench is a benchmarking framework designed to evaluate large language models' ability to generate optimized GPU kernels for multi-GPU computing environments. The benchmark assesses LLMs

A basic sigmoid extrapolation suggests that this measurement will be close to its performance ceiling in around a year---meaning most random…

SafetyDGX agent

A basic sigmoid extrapolation suggests that this measurement will be close to its performance ceiling in around a year---meaning most randomly sampled remote work projects would be highly automatable

A Bayesian Filtering Approach for Learning Lagrangian Dynamics from Noisy Measurements

ResearchDGX agent

arXiv:2606.31137v1 Announce Type: new Abstract: This paper proposes a Bayesian filtering-based approach for learning the dynamics of a physical system from partial, noisy measurements. We model the sy

a bit of a teaser for the upcoming talk @aiDotEngineer World's Fair 12:05 pm, Room 2003

ToolsDGX agent

a bit of a teaser for the upcoming talk @aiDotEngineer World's Fair 12:05 pm, Room 2003 Intelligence != Expertise Intelligence + Continual Learning = Expertise I will share some thoughts about continu

A Coherence Law for Trainability in Noisy Equivariant Quantum Neural Networks

ResearchDGX agent

arXiv:2606.30688v1 Announce Type: cross Abstract: Symmetry provides a quantum neural network structure, but on its own it does not keep the network trainable once noise is present. We ask which physic

A good little EV you won't be able to buy soon: The Volvo EX30 Cross Country

IndustryDGX agent

The 2026 Volvo EX30 Cross Country is a small electric SUV that is a joy to drive with a stylish cabin, though the rugged Cross Country variant sacrifices only modest range compared to the standard mod

A Large-Language-Model Supported Personalized Driving Framework for Lane Change in Highway Scenarios

Model ReleasesDGX agent

arXiv:2606.31483v1 Announce Type: new Abstract: Personalized driving can improve the user acceptance of automated driving systems. However, existing methods still provide limited support for translati

A Lifecycle and Application-Stack Survey of Large Language Model Vulnerabilities: Attacks, Risks, Defenses, and Open Problems

SafetyDGX agent

arXiv:2606.31639v1 Announce Type: cross Abstract: Large language models are no longer only text generators. They are increasingly embedded in retrieval pipelines, enterprise assistants, coding environ

A look at Manifest, an annual convention for prediction markets where many purists worry that Kalshi and Polymarket are undermining the technology's public good (Christopher Beam/Bloomberg)

IndustryDGX agent

Christopher Beam / Bloomberg: A look at Manifest, an annual convention for prediction markets where many purists worry that Kalshi and Polymarket are undermining the technology's public good — At Mani

A Modular Vision-Language-Action Robotics Framework for Indoor Environments

AgentsDGX agent

arXiv:2606.31144v1 Announce Type: cross Abstract: This paper presents an integrated system for the CMU Vision-Language-Action (VLA) Challenge, designed to enable an autonomous agent to perform complex

A once in a lifetime crossover between @aiDotEngineer and @FIFAWorldCup is happening in SF! 100 speakers from AIE got the golden pass to see…

ToolsDGX agent

A once in a lifetime crossover between @aiDotEngineer and @FIFAWorldCup is happening in SF! 100 speakers from AIE got the golden pass to see team USA kick Bosnia's ass!! Thanks @swyx @liamcbride and T

A Realistic Protocol for Evaluation of Weakly Supervised Object Localization

Model ReleasesDGX agent

arXiv:2404.10034v3 Announce Type: replace Abstract: Weakly Supervised Object Localization (WSOL) allows training deep learning models for classification and localization (LOC) using only global class-

A Reproducible Benchmark of Lightweight CNNs: Accuracy, Efficiency, and the Impact of Pretrained Initialization

Model ReleasesDGX agent

arXiv:2505.03303v3 Announce Type: replace-cross Abstract: Lightweight convolutional neural networks are often compared using results obtained with different training recipes, input settings, and pretr

A researcher says a vulnerability in Apple's Hide My Email tool lets anyone see a user's real email address; first reported in June 2025, Apple has not fixed it (Joseph Cox/404 Media)

IndustryDGX agent

Joseph Cox / 404 Media: A researcher says a vulnerability in Apple's Hide My Email tool lets anyone see a user's real email address; first reported in June 2025, Apple has not fixed it — “Hide My Emai

A Scalable Whole-body Motion Transfer via Implicit Kinodynamic Motion Retargeting

SafetyDGX agent

arXiv:2509.15443v2 Announce Type: replace-cross Abstract: Human-to-humanoid imitation learning presents a promising pathway to address the severe data scarcity bottleneck in robotics by utilizing abun

A Self-Evolving Agentic System for Automated Generation and Execution of Biological Protocols

Model ReleasesDGX agent

arXiv:2606.31763v1 Announce Type: new Abstract: Autonomous wet-lab experimentation requires more than plausible protocol text: biological intent, quantitative procedures, device constraints and experi

A Semantic-Layer-Mediated Agent for Natural Language to SQL over Heterogeneous Enterprise Databases

Model ReleasesDGX agent

arXiv:2606.31041v1 Announce Type: new Abstract: Natural language-to-SQL (NL2SQL) over real-world enterprise databases remains significantly more challenging than on academic benchmarks. Enterprise sch

A Single Rewrite Suffices: Empirical Lessons from Production Skill Description Optimization

AgentsDGX agent

arXiv:2606.30775v1 Announce Type: cross Abstract: Enterprise AI agents route user queries to specialized skills by matching queries against natural language skill descriptions. When two skills share o

A space history mystery: What happened to the Viking arm used 50 years ago?

IndustryDGX agent

The Viking 1 lander used its robotic arm to perform the first Martian soil sample analysis , collecting samples for biological experiments that involved scooping soil and testing it with nutrient solu

A Stationary-Distribution Theory for Triplet-Based Plateau Search in Random Forest Ensemble-Size Selection

Model ReleasesDGX agent

arXiv:2606.30837v1 Announce Type: cross Abstract: The number of trees is a central computational parameter in Random Forests: increasing it reduces finite-ensemble variability but increases training a

A swap-adversarial framework for improving domain generalization in electrocorticography-based Parkinson's disease classification

Model ReleasesDGX agent

arXiv:2602.10528v2 Announce Type: replace-cross Abstract: We propose a novel swap-adversarial framework that mitigates high inter-subject variability and the high-dimensional low-sample-size problem i

A Systematic Approach to Multi-Agent AI from Advanced Regulatory Control Theory: Safe and Auditable LLM Operator Agents for Process Control

Model ReleasesDGX agent

arXiv:2606.30877v1 Announce Type: cross Abstract: Recent literature shows that large language models (LLMs) are useful for general-purpose tasks yet perform poorly on specific domain ones. One reason

A Technical Typology of AI Systems in Public Administration

AgentsDGX agent

arXiv:2606.31755v1 Announce Type: cross Abstract: Research on artificial intelligence (AI) in the public sector often treats 'AI' as a single category, neglecting technical distinctions between differ

A Theory of How Pretraining Shapes Inductive Bias in Fine-Tuning

SafetyDGX agent

arXiv:2602.20062v2 Announce Type: replace Abstract: Pretraining and fine-tuning are central stages in modern machine learning systems. In practice, feature learning plays an important role across both

A Three-Phase Foundation Model for Tax-Aware Personalized Portfolio Management

Model ReleasesDGX agent

arXiv:2606.30997v1 Announce Type: new Abstract: We present a three-phase deep reinforcement learning system for personalized portfolio management that addresses three limitations shared by all prior f

A time-series classification framework for individual-level absenteeism prediction under severe class imbalance

Model ReleasesDGX agent

arXiv:2606.31532v1 Announce Type: new Abstract: Staff absenteeism imposes substantial operational costs in high-demand work environments such as healthcare, emergency services, meat processing, constr

A Transferable Learned Temporal Prior for Transmission Reconstruction and Decision-Relevant Uncertainty in Real Outbreak Labels

Model ReleasesDGX agent

arXiv:2606.30842v1 Announce Type: new Abstract: Outbreak transmission reconstruction treats epidemiological timing and transmission labels as deterministic ground truth; neither has been systematicall

← Previous
1…408409410411412…1464
Next →