📺 M5 Ultra… Apple Wasn’t Messing Around
I compare the M5 Ultra and M3 Ultra on local LLM workloads, focusing on prompt processing and token generation. The DeepSeek V4 Flash test also shows how prompt length affects the performance difference.
- Local LLM comparison: prompt processing and token generation
- How results vary with prompt length, from short prompts to longer inputs
- Scope: these two performance measures, rather than a broader system or workload comparison
If you use local LLMs for coding agents or other professional tasks, this comparison can help you assess the upgrade against your typical prompt lengths. Consider your workload before deciding whether the performance gains matter to you.
この動画を紹介した Alex Ziskind の最新動画も、紹介付きで読めます。
📄 このページの紹介文は AI が独自に生成したものであり、著作権をはじめとする他者の権利(商標権・名誉権・プライバシー等)を侵害しないよう配慮しています。動画の著作権は各作成者に帰属します。