AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

86,457Total entries
1Added by human
86,456Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,039 results
22 Apr 2026

Policy Gradient Primal-Dual Method for Safe Reinforcement Learning from Human Feedback

SafetyDGX agent

arXiv:2604.19024v1 Announce Type: new Abstract: Safe Reinforcement Learning from Human Feedback (Safe RLHF) has recently achieved empirical success in developing helpful and harmless large language mo

Polymarket says perpetuals are coming to the platform and lets users sign up for early access; it hasn't specified whether crypto perpetual futures are included (Tanaya Macheel/CNBC)

Model ReleasesDGX agent

Tanaya Macheel / CNBC: Polymarket says perpetuals are coming to the platform and lets users sign up for early access; it hasn't specified whether crypto perpetual futures are included — Prediction mar

RARE: Redundancy-Aware Retrieval Evaluation Framework for High-Similarity Corpora

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

arXiv:2604.19047v1 Announce Type: cross Abstract: Existing QA benchmarks typically assume distinct documents with minimal overlap, yet real-world retrieval-augmented generation (RAG) systems operate o

Realistic Handwritten Multi-Digit Writer (MDW) Number Recognition Challenges

Model ReleasesDGX agent

arXiv:2512.00676v2 Announce Type: replace Abstract: Isolated digit classification has served as a motivating problem for decades of machine learning research. In real settings, numbers often occur as

Reasoning-Aware AIGC Detection via Alignment and Reinforcement

SafetyDGX agent

arXiv:2604.19172v1 Announce Type: new Abstract: The rapid advancement and widespread adoption of Large Language Models (LLMs) have elevated the need for reliable AI-generated content (AIGC) detection,

Resolving the Robustness-Precision Trade-off in Financial RAG through Hybrid Document-Routed Retrieval

Model ReleasesDGX agent

arXiv:2603.26815v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) systems for financial document question answering typically follow a chunk-based paradigm: documents are

Revisiting Catastrophic Forgetting in Continual Knowledge Graph Embedding

ResearchDGX agent

arXiv:2604.19401v1 Announce Type: cross Abstract: Knowledge Graph Embeddings (KGEs) support a wide range of downstream tasks over Knowledge Graphs (KGs). In practice, KGs evolve as new entities and fa

Revisiting RaBitQ and TurboQuant: A Symmetric Comparison of Methods, Theory, and Experiments

Model ReleasesDGX agent

arXiv:2604.19528v1 Announce Type: cross Abstract: This technical note revisits the relationship between RaBitQ and TurboQuant under a unified comparison framework. We compare the two methods in terms

RT @lqiao: 🔥 I'm thrilled to welcome George Hu as President of Fireworks AI! 🔥 When George and I first met a year ago, I said to him – we…

ToolsDGX agent

Fireworks AI announced the appointment of George Hu as President of the company. The announcement came from Fireworks AI's leadership, noting that Hu and the company's founder met approximately a year

Rubrik rolls out Cloud SQL cyber resilience and Gemini agent governance at Google Cloud Next

Model ReleasesDGX agent

Cloud data management and data security company Rubrik Inc. today announced a deepening of its partnership with Google Cloud with two new integrations that extend its reach into managed database prote

SAGE: Training-Free Semantic Evidence Composition for Edge-Cloud Inference under Hard Uplink Budgets

ResearchDGX agent

arXiv:2604.19623v1 Announce Type: cross Abstract: Edge-cloud hybrid inference offloads difficult inputs to a powerful remote model, but the uplink channel imposes hard per-request constraints on the n

Sentipolis: Emotion-Aware Agents for Social Simulations

ResearchDGX agent

arXiv:2601.18027v2 Announce Type: replace Abstract: LLM agents are increasingly used for social simulation, yet emotion is often treated as a transient cue, causing emotional amnesia and weak long-hor

Sherpa.ai Privacy-Preserving Multi-Party Entity Alignment without Intersection Disclosure for Noisy Identifiers

SafetyDGX agent

arXiv:2604.19219v1 Announce Type: cross Abstract: Federated Learning (FL) enables collaborative model training among multiple parties without centralizing raw data. There are two main paradigms in FL:

SMART-Ship: A Comprehensive Synchronized Multi-modal Aligned Remote Sensing Targets Dataset and Benchmark for Berthed Ships Analysis

Model ReleasesDGX agent

arXiv:2508.02384v2 Announce Type: replace Abstract: Given the limitations of satellite orbits and imaging conditions, multi-modal remote sensing (RS) data is crucial in enabling long-term earth observ

Sources: Tencent and Alibaba are in talks to invest in DeepSeek at a 20B+ valuation, partly benchmarked against Moonshot's pending round at an 18B valuation (The Information)

Model ReleasesDGX agent

The Information: Sources: Tencent and Alibaba are in talks to invest in DeepSeek at a 20B+ valuation, partly benchmarked against Moonshot's pending round at an 18B valuation — Chinese tech giants Tenc

Sources: xAI held talks in recent weeks with Mistral and Cursor about a potential three-way partnership; Mistral co-founder Devendra Chaplot joined xAI in March (Grace Kay/Business Insider)

Model ReleasesDGX agent

Grace Kay / Business Insider: Sources: xAI held talks in recent weeks with Mistral and Cursor about a potential three-way partnership; Mistral co-founder Devendra Chaplot joined xAI in March — - Elon

STAR-Teaming: A Strategy-Response Multiplex Network Approach to Automated LLM Red Teaming

SafetyDGX agent

arXiv:2604.18976v1 Announce Type: new Abstract: While Large Language Models (LLMs) are widely used, they remain susceptible to jailbreak prompts that can elicit harmful or inappropriate responses. Thi

StepFly: Agentic Troubleshooting Guide Automation for Incident Diagnosis

Model ReleasesDGX agent

arXiv:2510.10074v2 Announce Type: replace Abstract: Effective incident management in large-scale IT systems relies on troubleshooting guides (TSGs), but their manual execution is slow and error-prone.

TabEmb: Joint Semantic-Structure Embedding for Table Annotation

TutorialsDGX agent

arXiv:2604.18939v1 Announce Type: new Abstract: Table annotation is crucial for making web and enterprise tables usable in downstream NLP applications. Unlike textual data where learning semantically

TabXEval: Why this is a Bad Table? An eXhaustive Rubric for Table Evaluation

Model ReleasesDGX agent

arXiv:2505.22176v3 Announce Type: replace Abstract: Evaluating tables qualitatively and quantitatively poses a significant challenge, as standard metrics often overlook subtle structural and content-l

Temp-R1: A Unified Autonomous Agent for Complex Temporal KGQA via Reverse Curriculum Reinforcement Learning

Model ReleasesDGX agent

arXiv:2601.18296v2 Announce Type: replace-cross Abstract: Temporal Knowledge Graph Question Answering (TKGQA) is inherently challenging, as it requires sophisticated reasoning over dynamic facts with

Tencent launches an international beta for QClaw, its OpenClaw-based AI agent, and says the Chinese version, launched in March, reached over 1M users in 10 days (T. K. Lin/KrASIA)

Model ReleasesDGX agent

T. K. Lin / KrASIA: Tencent launches an international beta for QClaw, its OpenClaw-based AI agent, and says the Chinese version, launched in March, reached over 1M users in 10 days — OpenClaw's founde

Tesla FSD V14.3.2 has just started rolling out to early access users. There is something new in the release notes: Tesla has unified the mod…

IndustryDGX agent

Tesla FSD V14.3.2 has just started rolling out to early access users. There is something new in the release notes: Tesla has unified the model between Actually Smart Summon, FSD, and Robotaxi for more

The Download: introducing the 10 Things That Matter in AI Right Now

Model ReleasesDGX agent

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. Introducing: 10 Things That Matter in AI Right Now What actual

The Essence of Balance for Self-Improving Agents in Vision-and-Language Navigation

SafetyDGX agent

arXiv:2604.19064v1 Announce Type: new Abstract: In vision-and-language navigation (VLN), self-improvement from policy-induced experience, using only standard VLN action supervision, critically depends

The LLM Inference Trilemma: Throughput, Latency, Cost

IndustryDGX agent

This article examines the fundamental trade-offs in large language model inference operations, specifically the competing priorities of maximizing throughput, minimizing latency, and reducing costs. I

Thinking Before Matching: A Reinforcement Reasoning Paradigm Towards General Person Re-Identification

TutorialsDGX agent

arXiv:2604.19218v1 Announce Type: new Abstract: Learning identity-discriminative representations with multi-scene generality has become a critical objective in person re-identification (ReID). However

this looks like a website but it’s an interactive generative video 🫨

AgentsDGX agent

this looks like a website but it’s an interactive generative video 🫨 Imagine every pixel on your screen, streamed live directly from a model. No HTML, no layout engine, no code. Just exactly what you

This wasn't the case with previous image generators, but the LLM you select has a huge effect on GPT-imagegen-2 output. GPT-5.4 Thinking and…

Model ReleasesDGX agent

This wasn't the case with previous image generators, but the LLM you select has a huge effect on GPT-imagegen-2 output. GPT-5.4 Thinking and GPT-5.4 Pro will produce much better images, especially for

Time-Scale Coupling Between States and Parameters in Recurrent Neural Networks

Model ReleasesDGX agent

arXiv:2508.12121v5 Announce Type: replace Abstract: We show that gating mechanisms in recurrent neural networks (RNNs) induce lag-dependent and direction-dependent effective learning rates, even when

Today we’re introducing two big steps for health at OpenAI: - ChatGPT for Clinicians, a free version of ChatGPT designed for clinical work -…

Model ReleasesDGX agent

Today we’re introducing two big steps for health at OpenAI: - ChatGPT for Clinicians, a free version of ChatGPT designed for clinical work - HealthBench Professional, a new benchmark to evaluate real

Treehub launches with Tim Draper and Anne Wojcicki to back the next wave of AI health founders

Model ReleasesDGX agent

Treehub, a new Stanford University-adjacent residency program backed by the AI Health Fund, launched today, with billionaire investor Tim Draper and 23andMe Holding Co. founder Anne Wojcicki among the

Truly sorry for any confusion or frustration caused by unclear, misleading, or inappropriate rules in our moderation system and on our pages…

Model ReleasesDGX agent

Truly sorry for any confusion or frustration caused by unclear, misleading, or inappropriate rules in our moderation system and on our pages. OpenClaw, Hermes, and SillyTavern are now explicitly marke

VideoAgent: Personalized Synthesis of Scientific Videos

Model ReleasesDGX agent

arXiv:2509.11253v2 Announce Type: replace Abstract: The technical complexity of research papers often limits their reach, necessitating more accessible formats like scientific videos to disseminate ke

VLM Performance:Qwen3.6-27B is natively multimodal, supporting both vision-language thinking and non-thinking modes in a single unified chec…

Model ReleasesDGX agent

VLM Performance:Qwen3.6-27B is natively multimodal, supporting both vision-language thinking and non-thinking modes in a single unified checkpoint — the same as Qwen3.6-35B-A3B. It handles images and

Voice of India: A Large-Scale Benchmark for Real-World Speech Recognition in India

Model ReleasesDGX agent

arXiv:2604.19151v1 Announce Type: new Abstract: Existing Indic ASR benchmarks often use scripted, clean speech and leaderboard driven evaluation that encourages dataset specific overfitting. In additi

We just hit #1 on the @huggingface BrowseComp-Plus leaderboard. Best accuracy: 92.53%. Best recall: 88.79%. Lowest calibration error across …

Model ReleasesDGX agent

We just hit #1 on the @huggingface BrowseComp-Plus leaderboard. Best accuracy: 92.53%. Best recall: 88.79%. Lowest calibration error across all submissions. Built with @AI21Labs Maestro. https://huggi

... which keeps things confusing, since it raises an important new question https://x.com/simonw/status/2046798283700617267

Model ReleasesDGX agent

... which keeps things confusing, since it raises an important new question https://x.com/simonw/status/2046798283700617267 @TheAmolAvasare If I sign up for a new $20/month account today and roll the

Workspace agents can work across tools—pulling context from docs, email, chats, code, and systems, and taking approved actions like updating…

Model ReleasesDGX agent

Workspace agents can work across tools—pulling context from docs, email, chats, code, and systems, and taking approved actions like updating @Linear issues, creating docs, or sending messages. In @Sla

Wrote up Anthropic's self-own about Claude Code pricing from this afternoon on my blog - it turned out they'd reversed course just as I hit …

Model ReleasesDGX agent

Wrote up Anthropic's self-own about Claude Code pricing from this afternoon on my blog - it turned out they'd reversed course just as I hit publish, so I've tried to update it to reflect the current s

X launches Custom Timelines, a Grok-powered feature letting users pin any of over 75 topics to their home tab, in early access to Premium subscribers on iOS (Nikita Bier/@nikitabier)

Model ReleasesDGX agent

Nikita Bier / @nikitabier: X launches Custom Timelines, a Grok-powered feature letting users pin any of over 75 topics to their home tab, in early access to Premium subscribers on iOS — Ladies and gen

ZC-Swish: Stabilizing Deep BN-Free Networks for Edge and Micro-Batch Applications

Model ReleasesDGX agent

arXiv:2604.19453v1 Announce Type: new Abstract: Batch Normalization (BN) is a cornerstone of deep learning, yet it fundamentally breaks down in micro-batch regimes (e.g., 3D medical imaging) and non-I

21 Apr 2026

A Comparative Evaluation of Geometric Accuracy in NeRF and Gaussian Splatting

Model ReleasesDGX agent

arXiv:2604.18205v1 Announce Type: new Abstract: Recent advances in neural rendering have introduced numerous 3D scene representations. Although standard computer vision metrics evaluate the visual qua

A Goal Without a Plan Is Just a Wish: Efficient and Effective Global Planner Training for Long-Horizon Agent Tasks

AgentsDGX agent

arXiv:2510.05608v2 Announce Type: replace Abstract: Agents based on large language models (LLMs) struggle with brainless trial-and-error and generating hallucinatory actions due to a lack of global pl

A High-Accuracy Optical Music Recognition Method Based on Bottleneck Residual Convolutions

SafetyDGX agent

arXiv:2604.16446v1 Announce Type: new Abstract: Optical Music Recognition (OMR) aims to convert printed or handwritten music score images into editable symbolic representations. This paper presents an

A Note on TurboQuant and the Earlier DRIVE/EDEN Line of Work

Model ReleasesDGX agent

arXiv:2604.18555v1 Announce Type: new Abstract: This note clarifies the relationship between the recent TurboQuant work and the earlier DRIVE (NeurIPS 2021) and EDEN (ICML 2022) schemes. DRIVE is a 1-

A Practical Guide to LLM Fine Tuning

TutorialsDGX agent

This guide from Databricks covers practical techniques and best practices for fine-tuning large language models, including methodologies for adapting pre-trained LLMs to specific tasks and domains. It

A Quasi-Experimental Developer Study of Security Training in LLM-Assisted Web Application Development

SafetyDGX agent

arXiv:2604.17763v1 Announce Type: cross Abstract: This paper presents a controlled quasi-experimental developer study examining whether a layer-based security training package is associated with impro

A Real-World Grasping-in-Clutter Performance Evaluation Benchmark for Robotic Food Waste Sorting

Model ReleasesDGX agent

arXiv:2602.18835v2 Announce Type: replace Abstract: Food waste management is critical for sustainability, yet inorganic contaminants hinder recycling potential. Robotic automation accelerates sorting

A Survey of Spatial Memory Representations for Efficient Robot Navigation

Model ReleasesDGX agent

arXiv:2604.16482v1 Announce Type: new Abstract: As vision-based robots navigate larger environments, their spatial memory grows without bound, eventually exhausting computational resources, particular

Adversarial Arena: Crowdsourcing Data Generation through Interactive Competition

SafetyDGX agent

arXiv:2604.17803v1 Announce Type: cross Abstract: Post-training Large Language Models requires diverse, high-quality data which is rare and costly to obtain, especially in low resource domains and for

Adverse-to-the-eXtreme Panoptic Segmentation: URVIS 2026 Study and Benchmark

Model ReleasesDGX agent

arXiv:2604.16984v1 Announce Type: new Abstract: This paper presents the report of the URVIS 2026 challenge on adverse-to-extreme panoptic segmentation. As the first challenge of its kind, it attracted

After users reported Claude Code appeared to be removed from Anthropic's Pro plan, Anthropic says it's 'running a small test on ~2% of new prosumer signups' (Ed Zitron/Ed Zitron's Where's Your Ed At)

Model ReleasesDGX agent

Ed Zitron / Ed Zitron's Where's Your Ed At: After users reported Claude Code appeared to be removed from Anthropic's Pro plan, Anthropic says it's “running a small test on ~2% of new prosumer signups”

AI Data Transformation Guide for Data Engineers and Data Scientists

TutorialsDGX agent

This guide from Databricks covers data transformation techniques and best practices essential for preparing data for AI/ML projects, addressing workflows that both data engineers and data scientists e

Align Documents to Questions: Question-Oriented Document Rewriting for Retrieval-Augmented Generation

SafetyDGX agent

arXiv:2604.17325v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) enhances the factuality of Large Language Models (LLMs) by incorporating retrieved documents and/or generated conte

Althea: Human-AI Collaboration for Fact-Checking and Critical Reasoning

Model ReleasesDGX agent

arXiv:2602.11161v2 Announce Type: replace-cross Abstract: The web's information ecosystem demands fact-checking systems that are both scalable and epistemically trustworthy. Automated approaches offer

Arena Trends: Text-to-Image, Jan 2026 – Apr 2026 For most of the year, @GoogleDeepMind and @OpenAI traded the top spot within a tight margin…

Model ReleasesDGX agent

Arena Trends: Text-to-Image, Jan 2026 – Apr 2026 For most of the year, @GoogleDeepMind and @OpenAI traded the top spot within a tight margin - GPT-Image vs. Nano Banana - with the rest of the field cl

Aspect Ratios & Resolution in ChatGPT Images 2.0, demonstrated by @dibyayB

Model ReleasesDGX agent

ChatGPT Images 2.0 supports multiple aspect ratios and resolutions for image generation, allowing users greater flexibility in creating images tailored to different use cases and display formats. The

Automatic Slide Updating with User-Defined Dynamic Templates and Natural Language Instructions

Model ReleasesDGX agent

arXiv:2604.17894v1 Announce Type: new Abstract: Presentation slides are a primary medium for data-driven reporting, yet keeping complex, analytics-style decks up to date remains labor-intensive. Exist

Back to Repair: A Minimal Denoising Network for Time Series Anomaly Detection

Model ReleasesDGX agent

arXiv:2604.17388v1 Announce Type: new Abstract: We introduce JuRe (Just Repair), a minimal denoising network for time series anomaly detection that exposes a central finding: architectural complexity

← Previous
1…710711712713714…1034
Next →