AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,724 results
10 Apr 2026

GIFT: Group-Relative Implicit Fine-Tuning Integrates GRPO with DPO and UNA

SafetyDGX agent

arXiv:2510.23868v4 Announce Type: replace Abstract: This paper proposes extit{Group-relative Implicit Fine-Tuning (GIFT)}, a reinforcement learning framework for aligning large language models (LLMs

GLM 5.1 on AI Gateway

ToolsDGX agent

GLM 5.1 from Z.ai is now available on Vercel's AI Gateway with no markup and no separate provider account required. Designed for long-horizon autonomous tasks, it handles planning, execution, testi...

Great to see this on X. This is something I created during my PhD time 20 years ago.

ResearchDGX agent

Great to see this on X. This is something I created during my PhD time 20 years ago. This is a healing grid by Japanese artist Ryota Kanai. If you stare at the center, the irregularities start to heal

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

HAWK: Head Importance-Aware Visual Token Pruning in Multimodal Models

HardwareDGX agent

arXiv:2604.07812v1 Announce Type: new Abstract: In multimodal large language models (MLLMs), the surge of visual tokens significantly increases the inference time and computational overhead, making th

HiF-VLA: Hindsight, Insight and Foresight through Motion Representation for Vision-Language-Action Models

ApplicationsDGX agent

arXiv:2512.09928v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models have recently enabled robotic manipulation by grounding visual and linguistic cues into actions. However, most V

If I had a pound for every time a MAGA account bellowed these mad MAGA myths me, I could fund European defence myself. 1. 'Europe would be c…

ApplicationsDGX agent

If I had a pound for every time a MAGA account bellowed these mad MAGA myths me, I could fund European defence myself. 1. 'Europe would be conquered in a week without America' Europe has 1.7 million t

Iteratively Learning Muscle Memory for Legged Robots to Master Adaptive and High Precision Locomotion

ResearchDGX agent

arXiv:2507.13662v2 Announce Type: replace Abstract: This paper presents a scalable and adaptive control framework for legged robots that integrates Iterative Learning Control (ILC) with a biologically

llama.cpp now supports various small OCR models that can run on low-end devices. These models are small enough to run on GPU with 4GB VRAM, …

Model ReleasesDGX agent

llama.cpp now supports various small OCR models that can run on low-end devices. These models are small enough to run on GPU with 4GB VRAM, and some of them can even run on CPU with decent performance

LongSpec: Long-Context Lossless Speculative Decoding with Efficient Drafting and Verification

ResearchDGX agent

arXiv:2502.17421v4 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) can now process extremely long contexts, efficient inference over these extended inputs has become increasingl

LongWriter-Zero: Mastering Ultra-Long Text Generation via Reinforcement Learning

Model ReleasesDGX agent

arXiv:2506.18841v3 Announce Type: replace-cross Abstract: Ultra-long generation by large language models (LLMs) is a widely demanded scenario, yet it remains a significant challenge due to their maxim

Loved the keynote talk about this at @aiDotEngineer . As a Age of Empires fan, I might try it -- just for fun :D Don't think I am going to u…

ToolsDGX agent

Loved the keynote talk about this at @aiDotEngineer . As a Age of Empires fan, I might try it -- just for fun :D Don't think I am going to use it for day to day but surely for fun! AgentCraft v1 is li

LPM 1.0: Video-based Character Performance Model

Model ReleasesDGX agent

arXiv:2604.07823v1 Announce Type: new Abstract: Performance, the externalization of intent, emotion, and personality through visual, vocal, and temporal behavior, is what makes a character alive. Lear

MARCH: Evaluating the Intersection of Ambiguity Interpretation and Multi-hop Inference

Model ReleasesDGX agent

arXiv:2509.22750v3 Announce Type: replace Abstract: Real-world multi-hop QA is naturally linked with ambiguity, where a single query can trigger multiple reasoning paths that require independent resol

MiniMax M2.7 is live on AI Gateway

ToolsDGX agent

MiniMax M2.7 is now available on Vercel AI Gateway in two variants: standard and high-speed, accessible without requiring separate provider accounts. It represents a significant improvement over pr...

Mitigating Distribution Sharpening in Math RLVR via Distribution-Aligned Hint Synthesis and Backward Hint Annealing

Model ReleasesDGX agent

arXiv:2604.07747v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) can improve low-k reasoning accuracy while narrowing solution coverage on challenging math que

MorphDistill: Distilling Unified Morphological Knowledge from Pathology Foundation Models for Colorectal Cancer Survival Prediction

SafetyDGX agent

arXiv:2604.06390v1 Announce Type: cross Abstract: Background: Colorectal cancer (CRC) remains a leading cause of cancer-related mortality worldwide. Accurate survival prediction is essential for treat

Now your business analyst in HR, Finance, or Sales team can build a live app or dashboard, a slide deck that updates in real time, or anythi…

ApplicationsDGX agent

Now your business analyst in HR, Finance, or Sales team can build a live app or dashboard, a slide deck that updates in real time, or anything for that matter for your CEO. Combine the speed and power

Often discuss my three-level vision for opening GLM to the community: First, we focus on accessibility by lowering the barrier to entry and …

HardwareDGX agent

Often discuss my three-level vision for opening GLM to the community: First, we focus on accessibility by lowering the barrier to entry and removing unnecessary constraints so developers can truly exp

On the way to @aiDotEngineer Europe 2026 with @swyx and @altryne for day three of the number one AI conference!

ToolsDGX agent

AI Engineer Europe 2026 was a three-day technical conference held April 8–10, 2026, at the Queen Elizabeth II Centre in London, marking the AI Engineer series' first flagship European event. It br...

Optimizing Vercel Sandbox snapshots

ToolsDGX agent

Vercel Sandbox snapshots capture the complete filesystem state of a running sandbox — including installed packages and configured environments — allowing new sandboxes to be launched from that save...

Orion-Lite: Distilling LLM Reasoning into Efficient Vision-Only Driving Models

Model ReleasesDGX agent

arXiv:2604.08266v1 Announce Type: new Abstract: Leveraging the general world knowledge of Large Language Models (LLMs) holds significant promise for improving the ability of autonomous driving systems

PokeGym: A Visually-Driven Long-Horizon Benchmark for Vision-Language Models

Model ReleasesDGX agent

arXiv:2604.08340v1 Announce Type: new Abstract: While Vision-Language Models (VLMs) have achieved remarkable progress in static visual understanding, their deployment in complex 3D embodied environmen

Q-Zoom: Query-Aware Adaptive Perception for Efficient Multimodal Large Language Models

Local AiDGX agent

arXiv:2604.06912v1 Announce Type: cross Abstract: MLLMs require high-resolution visual inputs for fine-grained tasks like document understanding and dense scene perception. However, current global res

Qwen 3.6 Plus on AI Gateway

Model ReleasesDGX agent

Alibaba's Qwen 3.6 Plus is now available on Vercel's AI Gateway, accessible via a unified API with no additional provider accounts required. Compared to Qwen 3.5 Plus, this model adds stronger age...

Reading Recognition in the Wild

TutorialsDGX agent

arXiv:2505.24848v4 Announce Type: replace Abstract: To enable egocentric contextual AI in always-on smart glasses, it is crucial to be able to keep a record of the user's interactions with the world,

Reason-SVG: Enhancing Structured Reasoning for Vector Graphics Generation with Reinforcement Learning

SafetyDGX agent

arXiv:2505.24499v2 Announce Type: replace Abstract: Generating high-quality Scalable Vector Graphics (SVGs) is challenging for Large Language Models (LLMs), as it requires advanced reasoning for struc

Research with ChatGPT

TutorialsDGX agent

The OpenAI Academy 'Research with ChatGPT' tutorial covers how to use ChatGPT's two primary research tools: the Search the Web feature for retrieving live data and sources quickly, and Deep Researc...

Riemann-Bench: A Benchmark for Moonshot Mathematics

Model ReleasesDGX agent

arXiv:2604.06802v1 Announce Type: new Abstract: Recent AI systems have achieved gold-medal-level performance on the International Mathematical Olympiad, demonstrating remarkable proficiency at competi

SciFigDetect: A Benchmark for AI-Generated Scientific Figure Detection

Model ReleasesDGX agent

arXiv:2604.08211v1 Announce Type: new Abstract: Modern multimodal generators can now produce scientific figures at near-publishable quality, creating a new challenge for visual forensics and research

sciwrite-lint: Verification Infrastructure for the Age of Science Vibe-Writing

Model ReleasesDGX agent

arXiv:2604.08501v1 Announce Type: cross Abstract: Science currently offers two options for quality assurance, both inadequate. Journal gatekeeping claims to verify both integrity and contribution, but

SealQA: Raising the Bar for Reasoning in Search-Augmented Language Models

Model ReleasesDGX agent

arXiv:2506.01062v4 Announce Type: replace Abstract: We introduce SealQA, a new challenge benchmark for evaluating SEarch-Augmented Language models on fact-seeking questions where web search yields con

Self-Distilled RLVR

SafetyDGX agent

arXiv:2604.03128v2 Announce Type: replace Abstract: On-policy distillation (OPD) has become a popular training paradigm in the LLM community. This paradigm selects a larger model as the teacher to pro

Sell More, Play Less: Benchmarking LLM Realistic Selling Skill

Model ReleasesDGX agent

arXiv:2604.07054v2 Announce Type: replace Abstract: Sales dialogues require multi-turn, goal-directed persuasion under asymmetric incentives, which makes them a challenging setting for large language

SepSeq: A Training-Free Framework for Long Numerical Sequence Processing in LLMs

Local AiDGX agent

arXiv:2604.07737v1 Announce Type: new Abstract: While transformer-based Large Language Models (LLMs) theoretically support massive context windows, they suffer from severe performance degradation when

Since our initial announcement, we have started onboarding beta design partners. - Connect your Lakebase data - Build an application - Deplo…

ApplicationsDGX agent

Since our initial announcement, we have started onboarding beta design partners. - Connect your Lakebase data - Build an application - Deploy to your Databricks account Reach out to your Replit/Databr

Soft-Quantum Algorithms

SafetyDGX agent

arXiv:2604.06523v1 Announce Type: cross Abstract: Quantum operations on pure states can be fully represented by unitary matrices. Variational quantum circuits, also known as quantum neural networks, e

Speaker room listens in on @badlogicgames

ToolsDGX agent

The specific tweet at that URL (status/2042526363899359295) is not publicly accessible via search results, and the content of that exact post could not be retrieved. However, based on closely relat...

TeamLLM: A Human-Like Team-Oriented Collaboration Framework for Multi-Step Contextualized Tasks

Model ReleasesDGX agent

arXiv:2604.06765v1 Announce Type: cross Abstract: Recently, multi-Large Language Model (LLM) frameworks have been proposed to solve contextualized tasks. However, these frameworks do not explicitly em

Tensor-Efficient High-Dimensional Q-learning

Model ReleasesDGX agent

arXiv:2511.03595v2 Announce Type: replace Abstract: High-dimensional reinforcement learning(RL) faces challenges with complex calculations and low sample efficiency in large state-action spaces. Q-lea

The Information wrote that Meta consumes 60.2T tokens per month, or ~750M per employee. Meta is getting absolutely token-mogged by us curren…

HardwareDGX agent

The Information wrote that Meta consumes 60.2T tokens per month, or ~750M per employee. Meta is getting absolutely token-mogged by us currently. SemiAnalysis employees consume 1.86B tokens per month.

The most awesome AI conference @aiDotEngineer by @swyx - doing one of the keynotes for it was the coolest thing! Loved that they brought it …

ToolsDGX agent

The AI Engineer conference (@aiDotEngineer) is a community and conference series for software engineers building with AI , organized by swyx (@swyx) and his co-founder Ben. The AI Engineer Code S...

The pace at which useful things are shipping also seems to be accelerating. Model releases are coming faster, of course, but so are signific…

TutorialsDGX agent

The pace at which useful things are shipping also seems to be accelerating. Model releases are coming faster, of course, but so are significant application and enterprise products (especially from Ant

🚨🇫🇷🇪🇺 This is big: the French government and agencies are officially getting out of Windows & non-EU tech. Each ministry has to present…

ResearchDGX agent

🚨🇫🇷🇪🇺 This is big: the French government and agencies are officially getting out of Windows & non-EU tech. Each ministry has to present their exit plan before Autumn: collaboration tools, antivirus, A

Towards Effective Long Video Understanding of Multimodal Large Language Models via One-shot Clip Retrieval

Model ReleasesDGX agent

arXiv:2512.08410v2 Announce Type: replace Abstract: Due to excessive memory overhead, most Multimodal Large Language Models (MLLMs) can only process videos of limited frames. In this paper, we propose

Using skills

TutorialsDGX agent

In ChatGPT, a skill is a reusable, shareable workflow — defined by a `SKILL.md` file — that tells ChatGPT how to consistently execute a specific task without starting from scratch each time. Skills...

Video Parallel Scaling: Aggregating Diverse Frame Subsets for VideoLLMs

Model ReleasesDGX agent

arXiv:2509.08016v2 Announce Type: replace Abstract: Video Large Language Models (VideoLLMs) face a critical bottleneck: increasing the number of input frames to capture fine-grained temporal detail le

Vision-Language Navigation for Aerial Robots: Towards the Era of Large Language Models

Model ReleasesDGX agent

arXiv:2604.07705v1 Announce Type: new Abstract: Aerial vision-and-language navigation (Aerial VLN) aims to enable unmanned aerial vehicles (UAVs) to interpret natural language instructions and autonom

What Makes an Ideal Quote? Recommending 'Unexpected yet Rational' Quotations via Novelty

SafetyDGX agent

arXiv:2602.22220v2 Announce Type: replace-cross Abstract: Quotation recommendation aims to enrich writing by suggesting quotes that complement a given context, yet existing systems mostly optimize sur

What's Missing in Screen-to-Action? Towards a UI-in-the-Loop Paradigm for Multimodal GUI Reasoning

Model ReleasesDGX agent

arXiv:2604.06995v1 Announce Type: new Abstract: Existing Graphical User Interface (GUI) reasoning tasks remain challenging, particularly in UI understanding. Current methods typically rely on direct s

When to Trust Tools? Adaptive Tool Trust Calibration For Tool-Integrated Math Reasoning

Model ReleasesDGX agent

arXiv:2604.08281v1 Announce Type: new Abstract: Large reasoning models (LRMs) have achieved strong performance enhancement through scaling test time computation, but due to the inherent limitations of

Yet another illustration of why LLMs aren’t even close to being AGI.

Model ReleasesDGX agent

Yet another illustration of why LLMs aren’t even close to being AGI. The world’s best LLMs are still terrible at poker. We put each model into a 200bb heads-up NLHE match against GTO Wizard AI. The be

9 Apr 2026

After playing with it a bit, Meta's Muse Spark Thinking is fine so far, but really doesn't match the current Big Three models. It also is a …

Model ReleasesDGX agent

After playing with it a bit, Meta's Muse Spark Thinking is fine so far, but really doesn't match the current Big Three models. It also is a bit... weird. Like some strange language & tone, a little lo

Apiiro launches command-line interface to bring AI-native security into software development workflows

Model ReleasesDGX agent

Application security posture management company Apiiro Ltd. today announced the launch of a new command-line interface designed to bring application security directly into artificial intelligence-driv

asgi-gzip 0.3

ApplicationsDGX agent

Release: asgi-gzip 0.3 I ran into trouble deploying a new feature using SSE to a production Datasette instance, and it turned out that instance was using datasette-gzip which uses asgi-gzip which was

Claude Mythos and misguided open-weight fearmongering

Model ReleasesDGX agent

The *Interconnects.ai* article argues that fears around releasing an open-weight version of Claude Mythos are overstated, noting that closed frontier models still lead open-weight ones in robust, ...

Come chat with me to dive deeper into the technical weeds on good harness engineering in Westminster!

ToolsDGX agent

Ryan Lopopolo, a Member of Technical Staff at OpenAI, posted this tweet inviting people to discuss harness engineering with him in-person at Westminster — referencing his broader work on the topic....

CyberAgent moves faster with ChatGPT Enterprise and Codex

ApplicationsDGX agent

CyberAgent, a major Japanese internet and media company, adopted ChatGPT Enterprise and OpenAI's Codex to accelerate engineering and business workflows across its organization. By integrating these...

doors open on first @aidotengineer keynotes coming up on livestream - @cramforce - @RaiaHadsell - @_lopopolo - @steipete join us on YouTube!…

ToolsDGX agent

The AI Engineer World's Fair (AIEWF) held its keynote livestream on YouTube, featuring prominent speakers including Malte Ubl (@cramforce, CTO at Vercel), Raia Hadsell (VP of Research at Google) ,...

https://x.com/ben_golub/status/2042079804271313343

TutorialsDGX agent

The specific tweet at that URL (status ID 2042079804271313343) could not be directly retrieved. However, based on the available context about Ben Golub's recent X/Twitter activity around Refine.ink...

If you want an example of what this looks like in practice, check out the '/research-docs' skill I created for Claude Code https://x.com/jer…

Model ReleasesDGX agent

If you want an example of what this looks like in practice, check out the '/research-docs' skill I created for Claude Code https://x.com/jerryjliu0/status/2041564207750246904?s=20 I built a Claude Cod

← Previous
1…293294295296
Next →