Ai feed
The world’s top Ai news.
Live headlines from the labs and outlets shaping Ai — rewritten in the Yai editorial voice for factory operators. Browse by player, country or topic. Refreshes every 15 minutes; every card links back to the original source.
- 36Kr13 min agoYai edit
36Kr offers iPhone 17 Pro in 10-second mini-program lottery
Users can enter a lottery for an iPhone 17 Pro by quickly experiencing the 36Kr Enterprise Intelligence mini-program and subscribing to a company, requiring no payment or referrals.
🇨🇳 ChinaBusinessRead at 36Kr - 🕰 Timeline · Anthropic22 days ago
Anthropic Claude Opus 4.8: Latest Opus tier; further improvements to reliability and tool use
Claude Opus 4.8 is the latest Opus tier, bringing further improvements to reliability and tool use. It is best used for frontier reasoning, long agentic runs, and judgment-heavy tasks.
Anthropic🇺🇸 USAModelsRead on yaikh.com - 36Kr58 min agoYai edit
Why game developers use Xiaohongshu for promotion
Game publishers are increasingly turning to Xiaohongshu for game promotion, even with limited budgets, despite its unconventional reputation compared to platforms like Bilibili.
🇨🇳 ChinaBusinessRead at 36Kr
QbitAI (量子位)1 hr agoYai editAliHealth's Hydrogen Ion partners with top medical journals
AliHealth's Hydrogen Ion has secured content partnerships with three major medical journals: NEJM, JAMA, and BMJ, marking a significant domestic achievement.
Alibaba🇨🇳 ChinaBusinessRead at QbitAI (量子位)- 🕰 Timeline · Anthropic25 days ago
Anthropic Claude 4.5 (Sonnet · Haiku): Sonnet 4.5 refines coding; Haiku 4.5 nears Sonnet 4 quality
Claude 4.5 Sonnet refines coding capabilities, while Haiku 4.5 approaches Sonnet 4 quality. It is best used for high-volume production requiring Sonnet-4 level intelligence at Haiku's cost.
Anthropic🇺🇸 USAModelsRead on yaikh.com
QbitAI (量子位)2 hr agoYai editAlibaba Qoder launches new security features with dedicated engineers
Alibaba's Qoder now offers enhanced security capabilities, providing each user with a dedicated security engineer.
Alibaba🇨🇳 ChinaBusinessRead at QbitAI (量子位)- 36Kr4 hr agoYai edit
Olin Biology restarts Hong Kong IPO bid amid strong vaccine pipeline
Olin Biology, a leading Chinese vaccine company, has refiled for a Hong Kong IPO, driven by a strong pipeline of innovative superbug vaccines and steady tetanus vaccine growth.
🇨🇳 ChinaBusinessRead at 36Kr - 36Kr5 hr agoYai edit
Shenzhen storage company re-applies for Hong Kong IPO
Shenzhen-based storage chip design company, Xin Tianxia, has resubmitted its application for a Hong Kong IPO after its initial filing expired, following a significant profit increase in Q1.
🇨🇳 ChinaBusinessRead at 36Kr - 🕰 Timeline · xAI21 days ago
xAI Grok 4: Latest tier; positioned as a general-purpose frontier competitor
Grok 4 is the latest tier, positioned as a general-purpose frontier competitor. It is best used for Grok API workloads and X-embedded products.
xAI🇺🇸 USAModelsRead on yaikh.com - 36Kr5 hr agoYai edit
Changzhou Jiaxuan Intelligent seeks Hong Kong IPO despite global leadership
Changzhou Jiaxuan Intelligent, a global leader in its sector, has filed for an IPO in Hong Kong, indicating a need for capital despite its market position.
🇨🇳 ChinaBusinessRead at 36Kr
TechCrunch AI7 hr agoYai editAnthropic's $1.5B copyright settlement approved
A court has approved Anthropic's $1.5 billion copyright settlement, resolving one case but not the broader issue of using copyrighted material for AI model training.
Anthropic🇺🇸 USARegulationRead at TechCrunch AI
TechCrunch AI9 hr agoYai editTrump's latest AI czar resigns from CAISI role
The director position for the Center for AI Standards and Innovation (CAISI) sees another departure as Trump's latest AI czar resigns.
🇺🇸 USARegulationRead at TechCrunch AI- 🕰 Timeline · Anthropic27 days ago
Anthropic Claude 4 (Opus 4 · Sonnet 4): Major coding jump, longer autonomous tool use
Claude 4 delivered a major coding jump, longer autonomous tool use, and improved agent behavior. It is best used for multi-hour agent runs and complex codebases.
Anthropic🇺🇸 USAModelsRead on yaikh.com
TechCrunch AI10 hr agoYai editGoogle develops new AI chip for Gemini efficiency
Alphabet is reportedly working on a new AI chip specifically designed to enhance the efficiency of its Gemini models.
Google🇺🇸 USAHardwareModelsRead at TechCrunch AI
TechCrunch AI11 hr agoYai editAI's key protocol becomes easier to use
A crucial AI protocol is adopting a "stateless" approach to session IDs, simplifying its use, similar to how many standard websites operate.
ResearchRead at TechCrunch AI
TechCrunch AI12 hr agoYai editX relaunches rebuilt Android app globally
X has announced the global availability of its rebuilt Android application after a year-long development effort.
xAI🇺🇸 USABusinessRead at TechCrunch AI- 🕰 Timeline · Alibaba18 days ago
Alibaba Qwen 3: Frontier tier with strong multilingual coverage.
This release shipped a frontier tier with strong multilingual coverage. It is best for any deployment where Chinese-language quality matters.
Alibaba🇨🇳 ChinaModelsRead on yaikh.com
QbitAI (量子位)21 hr agoYai editWAIC 2026 concludes, highlighting AI 2.0 industry shift
The WAIC 2026 conference has concluded, showcasing the transition of AI 2.0 from technological breakthroughs to practical industrial applications.
🇨🇳 ChinaBusinessRead at QbitAI (量子位)
QbitAI (量子位)21 hr agoYai editAgentic Infrastructure emerges as AGI foundation
A common Agentic Infrastructure is surfacing as the foundational layer for the AGI era, supporting various model manufacturers.
AgentsModelsRead at QbitAI (量子位)- OpenAI22 hr agoYai edit
OpenAI shares lessons on long-horizon model safety
OpenAI discusses insights from deploying long-running AI models, detailing new safety risks, observed failures, and improved safeguards through iterative deployment.
OpenAI🇺🇸 USASafetyModelsRead at OpenAI - 🕰 Timeline · DeepSeek25 days ago
DeepSeek DeepSeek-V2: MoE architecture with striking cost efficiency
DeepSeek-V2 features an MoE architecture known for its striking cost efficiency. It is best used for high-volume inference tasks where cost is the primary concern.
DeepSeek🇨🇳 ChinaModelsRead on yaikh.com
QbitAI (量子位)22 hr agoYai editAI boosts profits for rehabilitation center in smaller city
A rehabilitation institution in a fourth-tier city reports a 40% profit increase after integrating AI, demonstrating AI's impact in human-centric industries.
🇨🇳 ChinaBusinessRead at QbitAI (量子位)
Wired AI2 days agoYai editPeriod trackers may be spying on users
Period tracking apps are likely collecting user data, raising privacy concerns amidst other cybersecurity incidents like Russian cyberattacks and DHS breaches.
SafetyRead at Wired AI
Wired AI2 days agoYai editGoogle's new Gemini rates impact AI usage quotas
Google has revised its Gemini usage quotas, potentially reducing the number of AI responses users receive compared to previous rates.
Google🇺🇸 USAModelsBusinessRead at Wired AI- 🕰 Timeline · Anthropic22 days ago
Anthropic Claude 3 (Haiku · Sonnet · Opus): Three tiers for different needs
Claude 3 shipped in three tiers: Haiku for speed and low cost, Sonnet for balance, and Opus for the hardest tasks. It is best used when you need to pick a specific point on the cost/quality curve.
Anthropic🇺🇸 USAModelsRead on yaikh.com
Wired AI2 days agoYai editPrompt injection attacks thwart AI hacking agents
“Context bombing” is effectively neutralizing malicious AI agents by tricking them into shutting down before they can cause harm.
SafetyRead at Wired AI- OpenAI3 days agoYai edit
OpenAI CFO introduces AI scorecard for ROI measurement
OpenAI CFO Sarah Friar presents a practical AI scorecard to evaluate return on investment through metrics like useful work, cost per successful task, dependability, and return on compute.
OpenAI🇺🇸 USABusinessRead at OpenAI - 🕰 Timeline · Alibaba28 days ago
Alibaba Qwen 1.5: Multi-tier lineup from 0.5B to 72B, all open weights.
This release shipped a multi-tier lineup from 0.5B to 72B, all open weights. It is best for multilingual production LLMs, especially APAC.
Alibaba🇨🇳 ChinaModelsRead on yaikh.com - 🕰 Timeline · Anthropic27 days ago
Anthropic Claude 5 (Fable · Sonnet): New family; Fable optimised for narrative and voice
Claude 5, a new family alongside Opus 4.7/4.8, includes Fable, which is optimized for narrative and voice. It is best used for creative writing, dialogue, and podcast production.
Anthropic🇺🇸 USAModelsRead on yaikh.com - 🕰 Timeline · Anthropic23 days ago
Anthropic Claude Opus 4.7: Sharper reasoning and steering; strong default for hard analytical work
Claude Opus 4.7 offers sharper reasoning and steering, making it a strong default for hard analytical work. It is best used for deep research, code architecture, and high-quality writing.
Anthropic🇺🇸 USAModelsRead on yaikh.com - 🕰 Timeline · OpenAI21 days ago
OpenAI GPT-5: OpenAI's frontier tier; higher reliability plus stronger tool use
GPT-5 represents OpenAI's frontier tier, offering higher reliability and stronger tool use. It is best used for agentic workflows, long tool chains, and high-stakes decision support.
OpenAI🇺🇸 USAModelsRead on yaikh.com - 🕰 Timeline · Google18 days ago
Google Gemini 2.5 (Flash · Pro): Thinking modes and improved coding; strong price/performance
Gemini 2.5 (Flash · Pro) introduced thinking modes and improved coding, offering strong price/performance. It is best used for production LLM backends, translation, and batch content pipelines.
Google🇺🇸 USAModelsRead on yaikh.com - 🕰 Timeline · Meta18 days ago
Meta Llama 4: Native multimodal and long context, still open weights
Llama 4 shipped with native multimodal and long context capabilities, while remaining open weights. It is best used for multimodal applications where end-to-end ownership is critical.
Meta🇺🇸 USAModelsRead on yaikh.com - 🕰 Timeline · xAI20 days ago
xAI Grok 3: Trained on the Colossus cluster; competitive on reasoning benchmarks
Grok 3 was trained on the Colossus cluster and is competitive on reasoning benchmarks. It is best used for reasoning tasks, though still tightly tied to the X ecosystem.
xAI🇺🇸 USAModelsRead on yaikh.com - 🕰 Timeline · Anthropic27 days ago
Anthropic Claude 3.7 Sonnet: Extended thinking mode; toggleable deliberation for hard problems
Claude 3.7 Sonnet shipped with an extended thinking mode, offering toggleable deliberation for hard problems. It is best used for adaptive workloads, providing quick responses for easy tasks and deep analysis for complex ones.
Anthropic🇺🇸 USAModelsRead on yaikh.com - 🕰 Timeline · DeepSeek21 days ago
DeepSeek DeepSeek-R1: Open-weights reasoning model that rattled Silicon Valley.
This release shipped an open-weights reasoning model that rattled Silicon Valley. It is best for math, code, and reasoning tasks with self-hosted control.
DeepSeek🇨🇳 ChinaModelsRead on yaikh.com - 🕰 Timeline · Google25 days ago
Google Gemini 2.0: Native agentic tool use and stronger multi-step reasoning
Gemini 2.0 shipped with native agentic tool use and stronger multi-step reasoning capabilities. It is best used for agentic browser tasks and Workspace copilots.
Google🇺🇸 USAModelsRead on yaikh.com - 🕰 Timeline · DeepSeek24 days ago
DeepSeek DeepSeek-V3: Frontier-class quality trained at a fraction of Western costs.
This release shipped frontier-class quality trained at a fraction of Western costs. It is best for cost-sensitive assistants and as a reference for efficient training.
DeepSeek🇨🇳 ChinaModelsRead on yaikh.com - 🕰 Timeline · Alibaba29 days ago
Alibaba Qwen 2.5 (+ Coder, Math): Family expansion with specialist coding and math variants.
This release shipped a family expansion with specialist coding and math variants. It is best for coding copilots, math tutors, and structured extraction.
Alibaba🇨🇳 ChinaModelsRead on yaikh.com - 🕰 Timeline · OpenAI22 days ago
OpenAI o1 (reasoning): Chain-of-thought at the model level; trades latency for correctness
o1 (reasoning) shipped with chain-of-thought at the model level, trading latency for correctness. It is best used for math, hard coding, and formal reasoning, especially when waiting for an accurate answer is acceptable.
OpenAI🇺🇸 USAModelsRead on yaikh.com - 🕰 Timeline · xAI22 days ago
xAI Grok 2: Bigger jump in reasoning; image generation via Flux
Grok 2 delivered a bigger jump in reasoning capabilities and introduced image generation via Flux. It is best used for multimodal X integrations.
xAI🇺🇸 USAModelsRead on yaikh.com - 🕰 Timeline · Meta20 days ago
Meta Llama 3.1 (405B): First truly frontier-class open-weights model
Llama 3.1 (405B) was the first truly frontier-class open-weights model. It is best used for serious open-weight research and in regulated industries.
Meta🇺🇸 USAModelsRead on yaikh.com - 🕰 Timeline · Mistral18 days ago
Mistral Mistral Large 2 / Nemo: Updated frontier tier plus a compact multilingual model
Mistral Large 2 / Nemo includes an updated frontier tier and a compact multilingual model. It is best used for multilingual production LLMs on European infrastructure.
Mistral🇫🇷 FranceModelsRead on yaikh.com - 🕰 Timeline · Anthropic20 days ago
Anthropic Claude 3.5 Sonnet: Beat Claude 3 Opus on coding at Sonnet cost; the workhorse era
Claude 3.5 Sonnet surpassed Claude 3 Opus in coding performance at a Sonnet-level cost, ushering in the workhorse era. It is best used for production coding assistants and agent backbones.
Anthropic🇺🇸 USAModelsRead on yaikh.com - 🕰 Timeline · Alibaba26 days ago
Alibaba Qwen 2: Substantial quality jump; competitive with Llama 3 tier.
This release shipped a substantial quality jump, competitive with the Llama 3 tier. It is best for fine-tuning bases for enterprise Chinese/English apps.
Alibaba🇨🇳 ChinaModelsRead on yaikh.com - 🕰 Timeline · OpenAI23 days ago
OpenAI GPT-4o: Native multimodal (text, vision, audio) with real-time voice
GPT-4o shipped with native multimodal capabilities across text, vision, and audio, including real-time voice. It is best used for voice interfaces, live translation, and mixed-media assistants.
OpenAI🇺🇸 USAModelsRead on yaikh.com - 🕰 Timeline · Google22 days ago
Google Gemini 1.5 Flash: Cheap, fast tier of the 1.5 family for high-volume workloads
Gemini 1.5 Flash is a cheap, fast tier within the 1.5 family, designed for high-volume workloads. It is best used for real-time applications, feed rewrites, and batch classification.
Google🇺🇸 USAModelsRead on yaikh.com - 🕰 Timeline · Mistral27 days ago
Mistral Codestral: Code-specialised model for autocomplete and refactor
Codestral is a code-specialized model designed for autocomplete and refactoring. It is best used for IDE assistants and self-hosted coding copilots.
Mistral🇫🇷 FranceModelsRead on yaikh.com - 🕰 Timeline · Meta28 days ago
Meta Llama 3: 8B and 70B tiers; approaches GPT-4 quality at self-hosted cost
Llama 3 shipped in 8B and 70B tiers, approaching GPT-4 quality at a self-hosted cost. It is best used for on-premise enterprise LLMs and as a base for fine-tuning.
Meta🇺🇸 USAModelsRead on yaikh.com - 🕰 Timeline · Google27 days ago
Google Gemini 1.5 Pro: 1M-token context — the long-context breakthrough
Gemini 1.5 Pro shipped with a 1M-token context, marking a significant breakthrough in long-context processing. It is best used for feeding entire books, hours of video, or full codebases in a single prompt.
Google🇺🇸 USAModelsRead on yaikh.com - 🕰 Timeline · Mistral24 days ago
Mistral Mistral Large: Closed frontier tier; European sovereignty pitch
Mistral Large is a closed frontier tier model, emphasizing a European sovereignty pitch. It is best used for EU-based enterprises with specific data-residency needs.
Mistral🇫🇷 FranceModelsRead on yaikh.com - 🕰 Timeline · Google25 days ago
Google Gemini 1.0: Google's native multimodal frontier; Ultra, Pro, Nano tiers
Gemini 1.0 was Google's native multimodal frontier model, available in Ultra, Pro, and Nano tiers. It is best used for multimodal reasoning within the Google ecosystem.
Google🇺🇸 USAModelsRead on yaikh.com - 🕰 Timeline · Mistral24 days ago
Mistral Mixtral 8x7B: Mixture-of-experts open weights; strong quality for the price
Mixtral 8x7B is a mixture-of-experts open-weights model, offering strong quality for its price. It is best used for self-hosted assistants requiring GPT-3.5-class quality.
Mistral🇫🇷 FranceModelsRead on yaikh.com - 🕰 Timeline · xAI26 days ago
xAI Grok 1: Elon Musk's first release; snarky persona, integrated with X
Grok 1 was Elon Musk's first release, featuring a snarky persona and integration with X. It is best used for X-native content and personality-driven chat interactions.
xAI🇺🇸 USAModelsRead on yaikh.com - 🕰 Timeline · Anthropic23 days ago
Anthropic Claude 2.1: 200K-token context; the first model that ate whole codebases in one shot
Claude 2.1 shipped with a 200K-token context, making it the first model capable of processing entire codebases in one go. It is best used for whole-repo code review and long-document Q&A.
Anthropic🇺🇸 USAModelsRead on yaikh.com - 🕰 Timeline · Mistral26 days ago
Mistral Mistral 7B: Open-weights 7B model that punched above its size
Mistral 7B was an open-weights 7B model that performed remarkably well for its size. It is best used for edge, laptop, and cost-constrained inference scenarios.
Mistral🇫🇷 FranceModelsRead on yaikh.com - 🕰 Timeline · Alibaba28 days ago
Alibaba Qwen 1: Alibaba's first open-weights Qwen models.
This release shipped Alibaba's first open-weights Qwen models. It is best for Chinese-language assistants and Alibaba Cloud deployments.
Alibaba🇨🇳 ChinaModelsRead on yaikh.com - 🕰 Timeline · Meta28 days ago
Meta Llama 2: First openly licensed model competitive with GPT-3.5
Llama 2 was the first openly licensed model that was competitive with GPT-3.5. It is best used for self-hosted assistants and privacy-sensitive deployments.
Meta🇺🇸 USAModelsRead on yaikh.com - 🕰 Timeline · Anthropic24 days ago
Anthropic Claude 2: 100K-token context and stronger coding; the enterprise breakout
Claude 2 shipped with a 100K-token context and stronger coding abilities, marking its enterprise breakout. It is best used for legal, research, and analyst workflows involving long documents.
Anthropic🇺🇸 USAModelsRead on yaikh.com - 🕰 Timeline · Anthropic18 days ago
Anthropic Claude 1: Anthropic's first public model — long context, strong helpfulness training
Claude 1 was Anthropic's first public model, featuring long context and strong helpfulness training. It is best used for long-document summarisation and safe assistants.
Anthropic🇺🇸 USAModelsRead on yaikh.com - 🕰 Timeline · OpenAI18 days ago
OpenAI GPT-4: First reliably 'smart' GPT — better reasoning, longer context, plus vision
GPT-4 was the first reliably 'smart' GPT, featuring better reasoning, longer context, and vision capabilities. It is best used for serious knowledge work, including analysis, coding, and structured extraction.
OpenAI🇺🇸 USAModelsRead on yaikh.com - 🕰 Timeline · Google19 days ago
Google Bard: Google's first ChatGPT response, initially built on LaMDA/PaLM
Bard was Google's initial response to ChatGPT, built on LaMDA/PaLM. It is best used for historical reference, as it was quickly superseded by Gemini.
Google🇺🇸 USAModelsRead on yaikh.com - 🕰 Timeline · Meta19 days ago
Meta LLaMA 1: Meta's first Llama — 'leaked' onto the open internet, changed everything
LLaMA 1 was Meta's first Llama model, which 'leaked' onto the open internet and fundamentally changed the Ai landscape. It is best used for kicking off the open-weights ecosystem.
Meta🇺🇸 USAModelsRead on yaikh.com - 🕰 Timeline · OpenAI18 days ago
OpenAI GPT-3.5: Powers the original ChatGPT free tier — fast, cheap, conversational
GPT-3.5 powered the original ChatGPT free tier, offering fast, cheap, and conversational capabilities. It is best used for chat assistants, drafts, and summaries where cost is prioritized over depth.
OpenAI🇺🇸 USAModelsRead on yaikh.com - 🕰 Timeline · OpenAI25 days ago
OpenAI GPT-3: 175B parameters; few-shot learning shocks the field
GPT-3, with 175 billion parameters, introduced few-shot learning that significantly impacted the field. It is best used for prompt-engineering-driven products, marking the era of 'just ask nicely'.
OpenAI🇺🇸 USAModelsRead on yaikh.com - 🕰 Timeline · OpenAI18 days ago
OpenAI GPT-2: 1.5B parameters, coherent long-form text, weights briefly withheld
GPT-2 shipped with 1.5 billion parameters, generating coherent long-form text, with OpenAI briefly withholding its weights. It is best used for long-form generation demos and sparked the first serious debates about Ai's potential dangers.
OpenAI🇺🇸 USAModelsRead on yaikh.com - 🕰 Timeline · OpenAI27 days ago
OpenAI GPT-1: Unsupervised pre-training + fine-tuning for NLP
This release proved out unsupervised pre-training with task-specific fine-tuning on NLP tasks. It is best used as a research foundation, showing that transformers scale, rather than for production.
OpenAI🇺🇸 USAModelsRead on yaikh.com - 📚 Yai History · EP116 days ago
Can a Machine Think? Turing's Question That Started It All
In 1950, Alan Turing asked if machines could think, proposing the Imitation Game to define intelligence by behavior. His foundational framing still shapes every debate about Ai's true intelligence today.
HistoryRead on yaikh.com - 📚 Yai History · EP217 days ago
Dartmouth Workshop: The Moment 'Artificial Intelligence' Was Coined
The 1956 Dartmouth workshop, led by John McCarthy, named the field 'Artificial Intelligence.' This early optimism, however, also set the stage for the first Ai winter two decades later.
HistoryRead on yaikh.com - 📚 Yai History · EP318 days ago
The Perceptron: Ai's First Learning Machine and Its Early Setback
Frank Rosenblatt's 1958 Perceptron was the first hardware neural net, but Minsky and Papert's 1969 critique halted neural network research for 15 years. This early crash shows why Ai's progress has often been stop-start rather than continuous.
HistoryRead on yaikh.com - 📚 Yai History · EP419 days ago
Expert Systems: Ai's First Commercial Wave Powered by Hand-Written Rules
The 1970s and 80s saw expert systems like DENDRAL and MYCIN achieve narrow success, marking Ai's first commercial boom. Their inability to generalize beyond specific rules ultimately led to the second Ai winter.
HistoryRead on yaikh.com - 📚 Yai History · EP520 days ago
The Ai Winters: How Funding Crashes Reshaped the Field
Two major funding collapses in the 70s and late 80s made 'Artificial Intelligence' a career-limiting term. These winters shaped how researchers approach ambition, making caution around 'AGI' a lasting legacy.
HistoryRead on yaikh.com - 📚 Yai History · EP620 days ago
The Statistical Turn: When Machine Learning Overtook Symbolic Ai
The 1990s and 2000s saw statistical methods like SVMs and random forests dominate practical Ai applications. Techniques from this era still power a significant portion of production Ai and competitive machine learning today.
HistoryRead on yaikh.com - 📚 Yai History · EP722 days ago
AlexNet Wins ImageNet: The Dawn of Deep Learning in 2012
In 2012, AlexNet dramatically reduced the ImageNet error rate using GPUs, signaling the arrival of deep learning. This event is the single most repeated inflection point in modern Ai, tracing back to every major Ai valuation today.
HistoryRead on yaikh.com - 📚 Yai History · EP823 days ago
Attention Is All You Need: The Transformer Paper That Changed Everything
Google Brain's 2017 Transformer paper introduced the architecture behind every modern Large Language Model. It is the closest thing the Ai field has to a moon-landing moment, fundamentally reshaping how models are built.
HistoryRead on yaikh.com - 📚 Yai History · EP924 days ago
The GPT Era Begins: Scaling Laws and Unexpected Capabilities
From 2018 to 2020, OpenAI's GPT series demonstrated the power of unsupervised pre-training and scaling laws. This era established the paradigm that making models larger often leads to surprising new capabilities.
HistoryRead on yaikh.com - 📚 Yai History · EP1025 days ago
ChatGPT: Ai's iPhone Moment in 2022
ChatGPT's launch in November 2022 made Ai accessible to millions, becoming the fastest-growing consumer product in history. This moment transformed Ai from a research topic into a critical boardroom discussion, leading to the creation of the Yai Ai feed.
HistoryRead on yaikh.com - 📚 Yai History · EP1126 days ago
Multi-Modal, Agentic, and Open: Ai Gets Loud From 2023
From 2023, Ai expanded into multi-modal capabilities, agentic tool use, and open-weight models like Llama. These advancements mean Ai is now something you deploy for practical tasks, making factory-floor use cases realistic.
HistoryRead on yaikh.com - 📚 Yai History · EP1226 days ago
The Frontier Today: Claude 5, GPT-5, Gemini 2.5
Today's state-of-the-art models, including Claude 5, GPT-5, and Gemini 2.5, represent the pinnacle of current Ai capabilities. Understanding this lineage helps you grasp the foundation of every model in your current Ai stack.
HistoryRead on yaikh.com