AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,323 results
22 Apr 2026

SAVOIR: Learning Social Savoir-Faire via Shapley-based Reward Attribution

Model ReleasesDGX agent

arXiv:2604.18982v1 Announce Type: new Abstract: Social intelligence, the ability to navigate complex interpersonal interactions, presents a fundamental challenge for language agents. Training such age

ShadowPEFT: Shadow Network for Parameter-Efficient Fine-Tuning

Model ReleasesDGX agent

arXiv:2604.19254v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning (PEFT) reduces the training cost of full-parameter fine-tuning for large language models (LLMs) by training only a sma

SitEmb-v1.5: Improved Context-Aware Dense Retrieval for Semantic Association and Long Story Comprehension

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2508.01959v2 Announce Type: replace Abstract: Retrieval-augmented generation (RAG) over long documents typically involves splitting the text into smaller chunks, which serve as the basic units f

Small and midsize businesses jumpstart their AI transformations with Gemini Enterprise

Model ReleasesDGX agent

Small businesses are the backbone of the global economy. With 400 million SMBs worldwide and 36 million in the U.S. alone, they provide 50% of global employment. Now, with Google Cloud AI, they’re sca

SMART-Ship: A Comprehensive Synchronized Multi-modal Aligned Remote Sensing Targets Dataset and Benchmark for Berthed Ships Analysis

Model ReleasesDGX agent

arXiv:2508.02384v2 Announce Type: replace Abstract: Given the limitations of satellite orbits and imaging conditions, multi-modal remote sensing (RS) data is crucial in enabling long-term earth observ

SmokeGS-R: Physics-Guided Pseudo-Clean 3DGS for Real-World Multi-View Smoke Restoration

Model ReleasesDGX agent

arXiv:2604.05301v2 Announce Type: replace Abstract: Real-world smoke simultaneously attenuates scene radiance, adds airlight, and destabilizes multi-view appearance consistency, making robust 3D recon

Sources: OpenAI has been briefing US federal agencies, state governments, and Five Eyes allies on the capabilities of its GPT-5.4-Cyber model over the past week (Sam Sabin/Axios)

Model ReleasesDGX agent

Sam Sabin / Axios: Sources: OpenAI has been briefing US federal agencies, state governments, and Five Eyes allies on the capabilities of its GPT-5.4-Cyber model over the past week — OpenAI has been br

Sources: Tencent and Alibaba are in talks to invest in DeepSeek at a 20B+ valuation, partly benchmarked against Moonshot's pending round at an 18B valuation (The Information)

Model ReleasesDGX agent

The Information: Sources: Tencent and Alibaba are in talks to invest in DeepSeek at a 20B+ valuation, partly benchmarked against Moonshot's pending round at an 18B valuation — Chinese tech giants Tenc

Sources: xAI held talks in recent weeks with Mistral and Cursor about a potential three-way partnership; Mistral co-founder Devendra Chaplot joined xAI in March (Grace Kay/Business Insider)

Model ReleasesDGX agent

Grace Kay / Business Insider: Sources: xAI held talks in recent weeks with Mistral and Cursor about a potential three-way partnership; Mistral co-founder Devendra Chaplot joined xAI in March — - Elon

SpecAgent: A Speculative Retrieval and Forecasting Agent for Code Completion

Model ReleasesDGX agent

arXiv:2510.17925v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) excel at code-related tasks but often struggle in realistic software repositories, where project-specific APIs an

SPRITE: From Static Mockups to Engine-Ready Game UI

Model ReleasesDGX agent

arXiv:2604.18591v1 Announce Type: cross Abstract: Game UI implementation requires translating stylized mockups into interactive engine entities. However, current 'Screenshot-to-Code' tools often strug

Startups are building the next big thing with Google Cloud AI

Model ReleasesDGX agent

The future is taking shape in Las Vegas this week, where the world’s leading startups are showcasing their groundbreaking AI work at Google Cloud Next. Whether they need the top AI models, infrastruct

StepFly: Agentic Troubleshooting Guide Automation for Incident Diagnosis

Model ReleasesDGX agent

arXiv:2510.10074v2 Announce Type: replace Abstract: Effective incident management in large-scale IT systems relies on troubleshooting guides (TSGs), but their manual execution is slow and error-prone.

Storage innovations to accelerate your AI workloads at Next ‘26

Model ReleasesDGX agent

At Google Cloud Next, we are announcing innovations across every layer of our storage stacks — performance, intelligence, and management — to ensure your data is as fast and as useful as the AI models

STReasoner: Empowering LLMs for Spatio-Temporal Reasoning in Time Series via Spatial-Aware Reinforcement Learning

Model ReleasesDGX agent

arXiv:2601.03248v2 Announce Type: replace Abstract: Spatio-temporal reasoning in time series involves the explicit synthesis of temporal dynamics, spatial dependencies, and textual context. This capab

TabReX : Tabular Referenceless eXplainable Evaluation

Model ReleasesDGX agent

arXiv:2512.15907v2 Announce Type: replace Abstract: Evaluating the quality of tables generated by large language models (LLMs) remains an open challenge: existing metrics either flatten tables into te

TabXEval: Why this is a Bad Table? An eXhaustive Rubric for Table Evaluation

Model ReleasesDGX agent

arXiv:2505.22176v3 Announce Type: replace Abstract: Evaluating tables qualitatively and quantitatively poses a significant challenge, as standard metrics often overlook subtle structural and content-l

Take Out Your Calculators: Estimating the Real Difficulty of Question Items with LLM Student Simulations

Model ReleasesDGX agent

arXiv:2601.09953v2 Announce Type: replace Abstract: Standardized math assessments require expensive human pilot studies to establish the difficulty of test items. We investigate the predictive value o

Talking to a Know-It-All GPT or a Second-Guesser Claude? How Repair reveals unreliable Multi-Turn Behavior in LLMs

Model ReleasesDGX agent

arXiv:2604.19245v1 Announce Type: cross Abstract: Repair, an important resource for resolving trouble in human-human conversation, remains underexplored in human-LLM interaction. In this study, we inv

Taming Actor-Observer Asymmetry in Agents via Dialectical Alignment

Model ReleasesDGX agent

arXiv:2604.19548v1 Announce Type: cross Abstract: Large Language Model agents have rapidly evolved from static text generators into dynamic systems capable of executing complex autonomous workflows. T

Temp-R1: A Unified Autonomous Agent for Complex Temporal KGQA via Reverse Curriculum Reinforcement Learning

Model ReleasesDGX agent

arXiv:2601.18296v2 Announce Type: replace-cross Abstract: Temporal Knowledge Graph Question Answering (TKGQA) is inherently challenging, as it requires sophisticated reasoning over dynamic facts with

Tencent launches an international beta for QClaw, its OpenClaw-based AI agent, and says the Chinese version, launched in March, reached over 1M users in 10 days (T. K. Lin/KrASIA)

Model ReleasesDGX agent

T. K. Lin / KrASIA: Tencent launches an international beta for QClaw, its OpenClaw-based AI agent, and says the Chinese version, launched in March, reached over 1M users in 10 days — OpenClaw's founde

The decisive layer in AI is still unclaimed: theCUBE’s Google Cloud Next day one keynote analysis

Model ReleasesDGX agent

The fight for the agent control plane is underway — and it might determine who controls enterprise artificial intelligence for the next decade. Google LLC came into Google Cloud Next 2026 with a clear

The Download: introducing the 10 Things That Matter in AI Right Now

Model ReleasesDGX agent

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. Introducing: 10 Things That Matter in AI Right Now What actual

The future of data lakehouse: Open and interoperable for the agentic era

Model ReleasesDGX agent

Traditional lakehouses were engineered for the era of reporting, not the high-velocity, multimodal demands of AI agents. To bridge this gap, architecture must evolve into an AI-native foundation — one

The High Explosives and Affected Targets (HEAT) Dataset

Model ReleasesDGX agent

arXiv:2604.18828v1 Announce Type: new Abstract: Artificial Intelligence (AI) surrogate models provide a computationally efficient alternative to full-physics simulations, but no public datasets curren

The new Gemini Enterprise: one platform for agent development, orchestration, and governance

Model ReleasesDGX agent

The first wave of AI changed how we find information; the next wave is changing how we get work done. Today, we’re enhancing our most powerful AI tools and bringing them together under one roof. Gemin

The Rise of Verbal Tics in Large Language Models: A Systematic Analysis Across Frontier Models

Model ReleasesDGX agent

arXiv:2604.19139v1 Announce Type: cross Abstract: As Large Language Models (LLMs) continue to evolve through alignment techniques such as Reinforcement Learning from Human Feedback (RLHF) and Constitu

This wasn't the case with previous image generators, but the LLM you select has a huge effect on GPT-imagegen-2 output. GPT-5.4 Thinking and…

Model ReleasesDGX agent

This wasn't the case with previous image generators, but the LLM you select has a huge effect on GPT-imagegen-2 output. GPT-5.4 Thinking and GPT-5.4 Pro will produce much better images, especially for

Time-Scale Coupling Between States and Parameters in Recurrent Neural Networks

Model ReleasesDGX agent

arXiv:2508.12121v5 Announce Type: replace Abstract: We show that gating mechanisms in recurrent neural networks (RNNs) induce lag-dependent and direction-dependent effective learning rates, even when

Time Series Augmented Generation for Financial Applications

Model ReleasesDGX agent

arXiv:2604.19633v1 Announce Type: new Abstract: Evaluating the reasoning capabilities of Large Language Models (LLMs) for complex, quantitative financial tasks is a critical and unsolved challenge. St

Today we’re introducing two big steps for health at OpenAI: - ChatGPT for Clinicians, a free version of ChatGPT designed for clinical work -…

Model ReleasesDGX agent

Today we’re introducing two big steps for health at OpenAI: - ChatGPT for Clinicians, a free version of ChatGPT designed for clinical work - HealthBench Professional, a new benchmark to evaluate real

Towards Optimal Agentic Architectures for Offensive Security Tasks

Model ReleasesDGX agent

arXiv:2604.18718v1 Announce Type: cross Abstract: Agentic security systems increasingly audit live targets with tool-using LLMs, but prior systems fix a single coordination topology, leaving unclear w

Towards Reliable Human Evaluations in Gesture Generation: Insights from a Community-Driven State-of-the-Art Benchmark

Model ReleasesDGX agent

arXiv:2511.01233v3 Announce Type: replace Abstract: We review human evaluation practices in automatic, speech-driven 3D gesture generation and find a lack of standardisation and frequent use of flawed

Towards Scalable Lifelong Knowledge Editing with Selective Knowledge Suppression

Model ReleasesDGX agent

arXiv:2604.19089v1 Announce Type: new Abstract: Large language models (LLMs) require frequent knowledge updates to reflect changing facts and mitigate hallucinations. To meet this demand, lifelong kno

Towards Understanding the Robustness of Sparse Autoencoders

Model ReleasesDGX agent

arXiv:2604.18756v1 Announce Type: cross Abstract: Large Language Models (LLMs) remain vulnerable to optimization-based jailbreak attacks that exploit internal gradient structure. While Sparse Autoenco

Treehub launches with Tim Draper and Anne Wojcicki to back the next wave of AI health founders

Model ReleasesDGX agent

Treehub, a new Stanford University-adjacent residency program backed by the AI Health Fund, launched today, with billionaire investor Tim Draper and 23andMe Holding Co. founder Anne Wojcicki among the

Truly sorry for any confusion or frustration caused by unclear, misleading, or inappropriate rules in our moderation system and on our pages…

Model ReleasesDGX agent

Truly sorry for any confusion or frustration caused by unclear, misleading, or inappropriate rules in our moderation system and on our pages. OpenClaw, Hermes, and SillyTavern are now explicitly marke

Tstars-Tryon 1.0: Robust and Realistic Virtual Try-On for Diverse Fashion Items

Model ReleasesDGX agent

arXiv:2604.19748v1 Announce Type: new Abstract: Recent advances in image generation and editing have opened new opportunities for virtual try-on. However, existing methods still struggle to meet compl

Two-dimensional early exit optimisation of LLM inference

Model ReleasesDGX agent

arXiv:2604.18592v1 Announce Type: cross Abstract: We introduce a two-dimensional (2D) early exit strategy that coordinates layer-wise and sentence-wise exiting for classification tasks in large langua

UniT: Toward a Unified Physical Language for Human-to-Humanoid Policy Learning and World Modeling

Model ReleasesDGX agent

arXiv:2604.19734v1 Announce Type: cross Abstract: Scaling humanoid foundation models is bottlenecked by the scarcity of robotic data. While massive egocentric human data offers a scalable alternative,

Unlocking the Edge deployment and ondevice acceleration of multi-LoRA enabled one-for-all foundational LLM

Model ReleasesDGX agent

arXiv:2604.18655v1 Announce Type: cross Abstract: Deploying large language models (LLMs) on smartphones poses significant engineering challenges due to stringent constraints on memory, latency, and ru

Unveiling Fine-Grained Visual Traces: Evaluating Multimodal Interleaved Reasoning Chains in Multimodal STEM Tasks

Model ReleasesDGX agent

arXiv:2604.19697v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have shown promising reasoning abilities, yet evaluating their performance in specialized domains remains chall

URoPE: Universal Relative Position Embedding across Geometric Spaces

Model ReleasesDGX agent

arXiv:2604.18747v1 Announce Type: new Abstract: Relative position embedding has become a standard mechanism for encoding positional information in Transformers. However, existing formulations are typi

VCE: A zero-cost hallucination mitigation method of LVLMs via visual contrastive editing

Model ReleasesDGX agent

arXiv:2604.19412v1 Announce Type: cross Abstract: Large vision-language models (LVLMs) frequently suffer from Object Hallucination (OH), wherein they generate descriptions containing objects that are

VDPP: Video Depth Post-Processing for Speed and Scalability

Model ReleasesDGX agent

arXiv:2604.06665v2 Announce Type: replace Abstract: Video depth estimation is essential for providing 3D scene structure in applications ranging from autonomous driving to mixed reality. Current end-t

VecHeart: Holistic Four-Chamber Cardiac Anatomy Modeling via Hybrid VecSets

Model ReleasesDGX agent

arXiv:2604.19403v1 Announce Type: new Abstract: Accurate cardiac anatomy modeling requires the model to be able to handle intricate interrelations among structures. In this paper, we propose VecHeart,

VideoAgent: Personalized Synthesis of Scientific Videos

Model ReleasesDGX agent

arXiv:2509.11253v2 Announce Type: replace Abstract: The technical complexity of research papers often limits their reach, necessitating more accessible formats like scientific videos to disseminate ke

ViDoRe V3: A Comprehensive Evaluation of Retrieval Augmented Generation in Complex Real-World Scenarios

Model ReleasesDGX agent

arXiv:2601.08620v2 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) pipelines must address challenges beyond simple single-document retrieval, such as interpreting visual elements

Visual Reasoning Agent: Robust Vision Systems in Remote Sensing via Inference-Time Scaling

Model ReleasesDGX agent

arXiv:2509.16343v2 Announce Type: replace-cross Abstract: Building robust vision systems for high-stakes domains such as remote sensing requires stronger visual reasoning than what single-pass inferen

Visual-TableQA: Open-Domain Benchmark for Reasoning over Table Images

Model ReleasesDGX agent

arXiv:2509.07966v2 Announce Type: replace-cross Abstract: Visual reasoning over structured data such as tables is a critical capability for modern vision-language models (VLMs), yet current benchmarks

VLA Foundry: A Unified Framework for Training Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2604.19728v1 Announce Type: cross Abstract: We present VLA Foundry, an open-source framework that unifies LLM, VLM, and VLA training in a single codebase. Most open-source VLA efforts specialize

VLM Performance:Qwen3.6-27B is natively multimodal, supporting both vision-language thinking and non-thinking modes in a single unified chec…

Model ReleasesDGX agent

VLM Performance:Qwen3.6-27B is natively multimodal, supporting both vision-language thinking and non-thinking modes in a single unified checkpoint — the same as Qwen3.6-35B-A3B. It handles images and

Voice of India: A Large-Scale Benchmark for Real-World Speech Recognition in India

Model ReleasesDGX agent

arXiv:2604.19151v1 Announce Type: new Abstract: Existing Indic ASR benchmarks often use scripted, clean speech and leaderboard driven evaluation that encourages dataset specific overfitting. In additi

Watch the Weights: Unsupervised monitoring and control of fine-tuned LLMs

Model ReleasesDGX agent

arXiv:2508.00161v3 Announce Type: replace-cross Abstract: The releases of powerful open-weight large language models (LLMs) are often not accompanied by access to their full training data. Existing in

We just hit #1 on the @huggingface BrowseComp-Plus leaderboard. Best accuracy: 92.53%. Best recall: 88.79%. Lowest calibration error across …

Model ReleasesDGX agent

We just hit #1 on the @huggingface BrowseComp-Plus leaderboard. Best accuracy: 92.53%. Best recall: 88.79%. Lowest calibration error across all submissions. Built with @AI21Labs Maestro. https://huggi

We need open traces so that everyone can train open agent models! cc @steipete @badlogicgames @thdxr @matanSF @hwchase17

Model ReleasesDGX agent

We need open traces so that everyone can train open agent models! cc @steipete @badlogicgames @thdxr @matanSF @hwchase17 People are misreading the SpaceX/Cursor deal as an M&A story. It’s actually a b

Welcome to the agentic BI era with Looker

Model ReleasesDGX agent

By combining the analytical depth of Looker with Google’s Agentic Data Cloud, the potential to transform how we model, interact with, and act on our data appears limitless. This week at Google Cloud N

We've published new research on how we post-train models for accurate search-augmented answers. Our SFT + RL pipeline improves search, citat…

Model ReleasesDGX agent

We've published new research on how we post-train models for accurate search-augmented answers. Our SFT + RL pipeline improves search, citation quality, instruction following, and efficiency. With Qwe

What’s new in BigQuery: Powering the Agentic Era

Model ReleasesDGX agent

Succeeding in the agentic era requires a transformation in your data strategy: moving from human-scale to agent-first workloads, evolving from reactive intelligence to proactive action, and shifting f

← Previous
1…323324325326327…373
Next →