[ THE_INFERENCE_HUB ]
Luma ↗Join the Hub

Where AI gets
shipped, not pitched.

The Inference Hub is a community of AI practitioners trading real production lessons: talks, events, and two active chats. Free to join, always.

Free forever. Read our Privacy Policy.
$ whoami# engineers shipping AI for a living$ ./join --free --no-vendor-pitches 2 chats connected (whatsapp, slack) next event loaded production lessons unlocked$

Where we've been

ALL_PAST_EVENTS
Are You Wasting Your Time With Harness Engineering?, Shuki Cohen × Gad Benram × Yaakov Tayeb · with AWS
MEETUP · Are You Wasting Your Time With Harness Engineering?
AI in Cybersecurity, Gad Benram × Guy Eshet × Niki Dvir · with Predict
MEETUP · AI in Cybersecurity
Darren Brian on AI, Darren Brian
TALK · Darren Brian on AI
AI for Data Science, Gad Benram × Miguel Neves
WEBINAR · AI for Data Science
Agent Scrum Tech, Tiago Santos
VIDEO_SERIES · Agent Scrum Tech
Emerging Architectures with LLMs 2025, Gabriel × Gad × Jeff
WEBINAR · Emerging Architectures with LLMs 2025
Levelling Up Your Operations in the Age of MCP, José Bastos
TALK · Levelling Up Your Operations in the Age of MCP
End-to-End TP(C)RM, Demi Ben-Ari
TALK · End-to-End TP(C)RM
Are You Wasting Your Time With Harness Engineering?, Shuki Cohen × Gad Benram × Yaakov Tayeb · with AWS
MEETUP · Are You Wasting Your Time With Harness Engineering?
AI in Cybersecurity, Gad Benram × Guy Eshet × Niki Dvir · with Predict
MEETUP · AI in Cybersecurity
Darren Brian on AI, Darren Brian
TALK · Darren Brian on AI
AI for Data Science, Gad Benram × Miguel Neves
WEBINAR · AI for Data Science
Agent Scrum Tech, Tiago Santos
VIDEO_SERIES · Agent Scrum Tech
Emerging Architectures with LLMs 2025, Gabriel × Gad × Jeff
WEBINAR · Emerging Architectures with LLMs 2025
Levelling Up Your Operations in the Age of MCP, José Bastos
TALK · Levelling Up Your Operations in the Age of MCP
End-to-End TP(C)RM, Demi Ben-Ari
TALK · End-to-End TP(C)RM
Are You Wasting Your Time With Harness Engineering?, Shuki Cohen × Gad Benram × Yaakov Tayeb · with AWS
MEETUP · Are You Wasting Your Time With Harness Engineering?
AI in Cybersecurity, Gad Benram × Guy Eshet × Niki Dvir · with Predict
MEETUP · AI in Cybersecurity
Darren Brian on AI, Darren Brian
TALK · Darren Brian on AI
AI for Data Science, Gad Benram × Miguel Neves
WEBINAR · AI for Data Science
Agent Scrum Tech, Tiago Santos
VIDEO_SERIES · Agent Scrum Tech
Emerging Architectures with LLMs 2025, Gabriel × Gad × Jeff
WEBINAR · Emerging Architectures with LLMs 2025
Levelling Up Your Operations in the Age of MCP, José Bastos
TALK · Levelling Up Your Operations in the Age of MCP
End-to-End TP(C)RM, Demi Ben-Ari
TALK · End-to-End TP(C)RM
JOIN_THE_CHAT

Get in the room

ON_THE_CALENDAR

Upcoming events

How do you build agents that stay on task for weeks, not minutes? A webinar on context management, memory, and architectures for long-horizon AI systems.

Utilization looks healthy and throughput still is not there. With Valar, we dig into measuring GPU inference properly: where workloads stall, which numbers matter, and how to read them.

Beyond the Click: AI in AdTech24 Sept 2026 · Online

How AdTech actually works, from real-time bidding to attribution, and what AI is really changing in production. No prior AdTech knowledge assumed. Live Q&A at the end.

ALL_EVENTS →
The Inference Hub logo
The Inference Hub

Follow us on Luma and new events land in your inbox the moment they go live.

Follow
WHO_WE_ARE

Built by practitioners

Tiago Santos speaking at a TensorOps event

The Inference Hub is a community of people who ship AI systems for a living. We meet to trade real production lessons: what breaks, what scales, and what it actually costs. No vendor pitches, no hype decks.

THE_ACADEMY

AI, explained properly

The Academy is our library of plain-language explainers: how this stuff actually works, from SVMs and XGBoost all the way to LLMs, RAG, and agents. Written by practitioners, free for everyone.

FIRST_POSTS_ARE_LIVE
What is an LLM?RAGTransformersEmbeddingsFine-tuningXGBoostSVMsNeural networksEvals
What is an LLM?RAGTransformersEmbeddingsFine-tuningXGBoostSVMsNeural networksEvals
What is an LLM?RAGTransformersEmbeddingsFine-tuningXGBoostSVMsNeural networksEvals
Vector databasesAgentsMCPQuantizationAttentionLoRAGradient boostingDistillationPrompt engineering
Vector databasesAgentsMCPQuantizationAttentionLoRAGradient boostingDistillationPrompt engineering
Vector databasesAgentsMCPQuantizationAttentionLoRAGradient boostingDistillationPrompt engineering
WHATS_INSIDE

What membership gets you

01Talks & meetups

Regular sessions with engineers shipping AI to production, online and in person.

02Two active chats

Daily chatter on WhatsApp, deep technical threads on Slack. Zero spam, no vendor pitches.

03Production lessons

What breaks, what scales, and what it actually costs, from people who've paid the bill.

SEE_YOU_INSIDE

The next event is already on the calendar.

One short form makes you a member. The next event invite lands in your inbox, and the chats are waiting.

Free forever. Read our Privacy Policy.