Jalapeño’s first results show industry-leading speed and efficiency in AI inference
Source
openai.com
Date
Terms in this piece · Glossary
inference — Running a trained model to get answers — the phase where AI is actually used, as opposed to trained.
Why it matters
OpenAI now has its own inferenceRunning a trained model to get answers — the phase where AI is actually used, as opposed to trained.Full definition → accelerator with published results, which bears on where serving capacity, latency and cost for its models are heading.