Home Glossary Inference

Inference - Page 7

Inference is the stage when a trained AI model uses new input to produce a prediction, classification, recommendation, or generated response. Training may take days across specialized hardware, while inference happens whenever a user submits a prompt, a camera captures an image, or an application requests a score. Production systems optimize inference for latency, throughput, memory use, energy consumption, and cost without degrading output quality. The model may run in a cloud service, data center, phone, vehicle, or edge device. Monitoring remains necessary because real-world inputs change, hardware can fail, and a model that performed well during evaluation may behave differently under live traffic.

Andrew Tulloch Leaves $12B AI Startup to Join Meta After Turning Down $1.5B Offer
By • 3 mins read
AI & Machine Learning, Immersive Reality (AR, VR, MR, and XR), News, Startups & Investment

Andrew Tulloch Leaves $12B AI Startup to Join Meta After Turning Down $1.5B Offer

By • 3 mins read

Andrew Tulloch, co-founder of the $12 billion AI startup Thinking Machines Lab, has joined Meta after previously rejecting what reports described as a $1.5 billion offer — a figure Meta has since called ‘inaccurate and ridiculous.’