A timeline turns contradictory instructions into an orderly sequence of updates. Without it the model just gambles between the old guideline and the new one on every run
As long as the relative time parser is hardcoded for japanese, english sessions are stuck with manual ISO dates. Wiring up dateparser or duckling would take an evening , so leaving that on the roadmap is an odd choice
Stories about Tim Cook sitting there being surprised by sales are total marketing. They simply ran out of memory chips because of the global shortage, so they repacked supply delays into a nice story about insane hype among AI startups. And hit two birds with one stone by throwing shade at their competitors
Apple hasn't been selling just ram for a long time, they sell vram. Try getting 512 gb of HBM on current Nvidia cards - it's gonna cost way more than $ 24k. And here you get the same amount of memory for weights right in a quiet unit under your desk
Put together a similar build with a couple of rtx 6000 Ada cards and Apple's price tag suddenly looks pretty damn reasonable
HBM itself is very expensive but it’s not really fair to compare to LPDDR or GDDR
They’re very different things.
The more logical argument to me is that Apple uses its upgrade price points as more than just direct BOM and rather as a proxy for things that are amortized across all their sales like support/warranty/etc so higher SKUs subsidize the costs of the lower ones.
If a kaiju shows up in the prompt, the weights will immediately drift from game theory into fiction. And by the laws of the genre, the military is obligated to drop a nuke on it - just to make the monster even angrier so it goes and trashes Tokyo
Yeah that’s why they delegated code gen to deterministic tooling and saved the model for input fuzzing - let it fight the data instead of legacy syntax
reply