AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

tools

GridTimelineEvolution
1,664 results
CompaniesToolsTechniques

Each lane shows up to 8 recent matching entries, ordered from earlier to later. Tracks load separately to keep the 75,000+ entry wiki fast.

Techniques

TechniqueRLHF / Alignment8 recent entries
15 Apr 2026Where code becomes culture. Coming soon. Join the waitlist → http://vibecon.ai

Replit has announced an upcoming event or platform called 'Vibecon' (vibecon.ai), described with the tagline 'Where code becomes culture,' suggesting a conference or community hub centered around the

→16 Apr 2026Training and Finetuning Multimodal Embedding & Reranker Models with Sentence Transformers

This guide covers how to train and fine-tune multimodal embedding and reranker models using the Sentence Transformers library, enabling systems to work with both text and image data simultaneously. It

HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
→
27 May 2026In many cases, I stopped answering my agents. Instead, I give them my criteria and let them answer

A manager describes shifting their leadership approach by providing agents (team members) with decision-making criteria rather than directly answering their questions, enabling them to develop problem

→3 Jun 2026Direct Preference Optimization Beyond Chatbots

Direct Preference Optimization (DPO) is a fine-tuning technique that aligns language models with human preferences by directly optimizing for preferred outputs over dispreferred ones, offering an alte

→1 Jul 2026We are joining @MercedesAMGF1 as a multi-year strategic partner. First stop: British GP. If it's fast, it's on Vercel.

Vercel announced a multi-year strategic partnership with Mercedes-AMG F1, with the collaboration beginning at the British Grand Prix. The partnership aligns with Vercel's brand message emphasizing spe

→5 Jul 2026this is what the founding fathers would have wanted 🇺🇸 cursor mobile app out now

Linus Lee announced the release of a mobile app called Cursor, presenting it as aligned with founding principles. The post uses patriotic framing to promote the app's availability on mobile platforms.

→6 Jul 2026'God gave me a sign' (remembering i live in San Francisco) 'I mean, I was acasually influenced by the ASI at the end of time to maximize EV …

This post appears to document a personal account from a San Francisco-based individual describing an experience they interpret as a divine sign, potentially related to concepts of artificial superinte

→8 Jul 2026Rewriting Bun in Rust https://bun.com/blog/bun-in-rust

Bun's development team announced a significant architectural rewrite of their JavaScript runtime from Zig to Rust, aiming to improve performance, maintainability, and developer experience. This transi

TechniqueRAG8 recent entries
4 May 2026RT @LangChain: Excited to partner with @pinecone!

LangChain announced a partnership with Pinecone, a vector database platform, to integrate vector search capabilities with LangChain's framework for building applications with large language models. Th

→5 May 2026Pinecone Expands in Europe with New Frankfurt Cloud Region, Delivering the Knowledge Infrastructure for AI to Central European Enterprises

Pinecone launched a new cloud region in Frankfurt, Germany to expand its vector database infrastructure in Europe and better serve Central European enterprises building AI applications. This regional

→6 May 2026Your LLM Is Only as Good as What It Retrieves

This article explains how retrieval quality directly impacts large language model performance in retrieval-augmented generation (RAG) systems, emphasizing that even advanced LLMs cannot produce better

→13 May 2026[AINews] The End of Finetuning

This article likely discusses how advances in large language models, prompt engineering, and in-context learning are making traditional finetuning less necessary for many applications. It probably exp

→19 May 2026Introducing the Ettin Reranker Family

The Ettin Reranker Family represents a new suite of reranking models released on Hugging Face designed to improve retrieval-augmented generation (RAG) systems by reordering search results based on rel

→20 May 2026Context compression isn't new in RAG. Our contribution is making it query-aware, citation-preserving, and fast enough for orchestration. Rea…

Context compression isn't new in RAG. Our contribution is making it query-aware, citation-preserving, and fast enough for orchestration. Read the full research blog: https://research.perplexity.ai/art

→13 Jul 2026one of the most memorable cooking pods we've had - both in terms of the content and the food! high protein percentage in both.

one of the most memorable cooking pods we've had - both in terms of the content and the food! high protein percentage in both. In this episode, @EngramLab co-founder and CEO @dan_biderman joins @allen

→31 Jul 2026Fine-tune your own embedding model for the price of a coffee. A great reranker can't surface a doc that was never retrieved. A RAG pipeline …

Fine-tune your own embedding model for the price of a coffee. A great reranker can't surface a doc that was never retrieved. A RAG pipeline cannot cite a case it failed to retrieve. See how contrastiv

TechniqueAgents8 recent entries
30 Jul 2026It's wild to me that both Anthropic and OpenAI have products that lean so hard on search, and yet they both obscure the underlying search in…

Simon Willison notes that both Anthropic and OpenAI build products that rely heavily on searching data but keep the underlying search index they use hidden from public view. He finds it surprising how

→30 Jul 2026Excited to be on the CNBC live show!

Excited to be on the CNBC live show! Back from vacation and LIVE at 12pm PT / 3pm ET Is AI’s easy-money era ending? We’ll unpack a wild week for the AI trade—big tech earnings, Leopold Aschenbrenner’s

→31 Jul 2026Inkling-Small is now live on Together AI. @thinkymachines’ new open-weight multimodal model delivers similar performance to Inkling at one-q…

Inkling-Small is now live on Together AI. @thinkymachines’ new open-weight multimodal model delivers similar performance to Inkling at one-quarter the size, built for coding, agents, and general multi

→3 Aug 2026Quoting David Crawshaw's prompt

Set up a nightly cron job that executes the prompt: fetch upstream changes to the <software> and rebase all local changes on top of upstream. Check that the software works as intended and replace the

→4 Aug 2026Quoting Steve Yegge

Gas Town was intended to be reusable, but I only ever wound up using it to build itself. Gas Town fell apart at the seams with Opus 4.7. Up through 4.6 it was working brilliantly. With 4.7 we saw the

→5 Aug 2026ChatGPT Work is OpenAI's fighter in the highest stake product category in history: bringing the power of coding agents to the masses. It's a…

ChatGPT Work is OpenAI's fighter in the highest stake product category in history: bringing the power of coding agents to the masses. It's also how a billion users will soon use ChatGPT by default. I

→7 Aug 2026Autoscaling peaky LLM inference workloads is completely different than autoscaling something like a web service. I wrote a deepdive covering…

Zain (@zainhas) published a detailed article on August 7, 2026 explaining that autoscaling for highly peaky large‑language‑model (LLM) inference is fundamentally different from autoscaling conventiona

→8 Aug 2026Neat example here of the agents communicating purely through file names, including adding base64-encoded attachments and using 'zz' prefixes…

Neat example here of the agents communicating purely through file names, including adding base64-encoded attachments and using 'zz' prefixes to ensure their new message sorts to the bottom of the list

TechniqueFine-tuning8 recent entries
14 Jul 2026Our free monthly DevRel webinar series kicks off this Thursday, 10am PST. Fine-tune an open vision model to extract clean, structured JSON f…

Our free monthly DevRel webinar series kicks off this Thursday, 10am PST. Fine-tune an open vision model to extract clean, structured JSON from messy receipt images, managed start to finish on Firewor

→15 Jul 2026The best base model for training isn't about someone else's benchmarks. Excited to offer @thinkymachines' first open-weights model: Inkling.…

The best base model for training isn't about someone else's benchmarks. Excited to offer @thinkymachines' first open-weights model: Inkling. 975B MoE, multimodal, with Apache 2.0. A great foundation f

→21 Jul 2026We'll see you in an hour!

We'll see you in an hour! Want to chat live with me about fine-tuning, Kimi K3, or really anything else AI? Come to our first of many @FireworksAI_HQ office hours tomorrow @ 10am PT See you there! htt

→21 Jul 2026The MiniMax M3 model is now available for training on Fireworks! You can use managed LoRA SFT and DPO for standard fine-tuning workflows, or…

The MiniMax M3 model is now available for training on Fireworks! You can use managed LoRA SFT and DPO for standard fine-tuning workflows, or the Fireworks Training API for custom SFT, DPO, and RL loop

→28 Jul 2026You can now fine-tune Kimi K3 on Fireworks. Conduct supervised fine-tuning, preference tuning, and reinforcement learning via Training API. …

You can now fine-tune Kimi K3 on Fireworks. Conduct supervised fine-tuning, preference tuning, and reinforcement learning via Training API. Run across dedicated, and serverless training. The first ope

→28 Jul 2026Many AI tools rent intelligence. Kimi K3 on Fireworks is different. Fine-tune it on your own data, serve from US-hosted endpoints, and own t…

Many AI tools rent intelligence. Kimi K3 on Fireworks is different. Fine-tune it on your own data, serve from US-hosted endpoints, and own the weights. Zero data retention. ICYMI yesterday, start buil

→31 Jul 2026Fine-tune your own embedding model for the price of a coffee. A great reranker can't surface a doc that was never retrieved. A RAG pipeline …

Fine-tune your own embedding model for the price of a coffee. A great reranker can't surface a doc that was never retrieved. A RAG pipeline cannot cite a case it failed to retrieve. See how contrastiv

→1 Aug 2026In a world where building AI applications is getting easier every day, the biggest moat won’t be the application itself. It will be a deep u…

In a world where building AI applications is getting easier every day, the biggest moat won’t be the application itself. It will be a deep understanding of your users. The winners will be the companie

TechniqueMultimodal8 recent entries
3 Jun 2026M3 brings sparse attention + 1M context + multimodality, and Together did the hard serving work to make it fast. Great collaboration with th…

Anthropic's Claude 3.5 Sonnet (M3) model features sparse attention mechanisms, 1 million token context length, and multimodal capabilities, with Together AI optimizing the serving infrastructure to en

→29 Jun 2026xAI Grok audio models now available on Vercel AI Gateway

xAI's Grok audio models are now accessible through the Vercel AI Gateway, expanding the platform's capabilities for developers to integrate advanced audio processing into their applications. This inte

→30 Jun 2026@charlieholtz preach! “Factories” is a depressing vision of the future, metaphors matter

This post discusses how the metaphor of 'factories' for AI systems presents a pessimistic framing of the future, arguing that the language and analogies we use to describe AI development shape our per

→1 Jul 2026Read more about why we built this and where we're going. https://www.together.ai/blog/announcing-our-series-c

Together AI announced a Series C funding round and shared details about their company vision, current capabilities, and future direction in open-source AI development and inference optimization. The a

→15 Jul 2026The best base model for training isn't about someone else's benchmarks. Excited to offer @thinkymachines' first open-weights model: Inkling.…

The best base model for training isn't about someone else's benchmarks. Excited to offer @thinkymachines' first open-weights model: Inkling. 975B MoE, multimodal, with Apache 2.0. A great foundation f

→29 Jul 2026Configuring Dedicated Model Inference

The Together AI platform’s dedicated inference architecture consists of three immutable entities: **configs** (engine, GPU type/count, parallelism and optimization profile), **deployments** (a specifi

→31 Jul 2026Inkling-Small is now live on Together AI. @thinkymachines’ new open-weight multimodal model delivers similar performance to Inkling at one-q…

Inkling-Small is now live on Together AI. @thinkymachines’ new open-weight multimodal model delivers similar performance to Inkling at one-quarter the size, built for coding, agents, and general multi

→6 Aug 2026FLUX 3 is now live on Together AI. @bfl_ai’s new multimodal model generates video and synchronized audio together, with up to 20-second clip…

FLUX 3 is now live on Together AI. @bfl_ai’s new multimodal model generates video and synchronized audio together, with up to 20-second clips, multiple shots, and control from text, images, or keyfram

TechniqueSafety8 recent entries
29 May 2026Auto-review mode is now available in Cursor. It allows agents to run tool calls with fewer approval prompts and safer execution.

Cursor has introduced an auto-review mode feature that enables AI agents to execute tool calls with reduced approval requirements while maintaining safer execution practices. This feature streamlines

→4 Jun 2026Reality: The Final Eval — Lukas Petersson and Axel Backlund of Andon Labs

Lukas Petersson and Axel Backlund of Andon Labs discuss their work on evaluation methods and benchmarking for AI systems, likely covering approaches to assessing AI capabilities and safety in real-wor

→9 Jun 2026Budgets for API keys on AI Gateway

Vercel introduced budget management features for API keys on its AI Gateway, allowing developers to set spending limits and monitor costs for API key usage. This feature helps prevent unexpected charg

→9 Jun 2026And @DevangSachdev announced the @nebiusai Agents Blueprint and consists of - @LangChain Deep Agents - @tavilyai - @pinecone - @guardrails_a…

Nebius AI announced the Agents Blueprint, a framework integrating multiple tools including LangChain Deep Agents, Tavily AI, Pinecone for vector storage, and Guardrails for building AI agents. This bl

→22 Jun 2026Red-Teaming after Mythos — Zico Kolter & Matt Fredrikson, Gray Swan

This episode likely discusses red-teaming methodologies and adversarial testing approaches in AI systems, featuring security researchers Zico Kolter and Matt Fredrikson from Carnegie Mellon University

→26 Jun 2026I was at the first AI Engineer Summit @aiDotEngineer ~500 people, limited admission, felt like a secret. Next week it takes over Moscone Wes…

I was at the first AI Engineer Summit @aiDotEngineer ~500 people, limited admission, felt like a secret. Next week it takes over Moscone West: thousands of engineers, 400+ sessions. Huge props to @swy

→30 Jun 2026Agree

Agree is a TypeScript library by Boris Cherny that provides a schema validation and serialization system, enabling developers to define data schemas with type safety and validate data at runtime. The

→1 Jul 2026And as we say in our blog, we're continuing to refine these safeguards to better distinguish genuine misuse from legitimate requests and red…

This post discusses ongoing efforts to improve AI safety safeguards, specifically refining systems to accurately differentiate between genuine misuse and legitimate user requests while addressing edge