|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|

Apple has been busy developing its own AI technologies, and recently offered a glimpse into how its models might work.
Currently, Apple’s direct competitors to the Meta Ray-Bans are planned for around 2027, together with AirPods equipped with cameras, which will provide their own set of AI-enabled capabilities.
While it’s still too early to anticipate what they will precisely look like, Apple unveiled MLX, its own open ML framework designed specifically for Apple Silicon.
Essentially, MLX provides a lightweight method to train and run models directly on Apple devices, remaining familiar to developers who prefer frameworks and languages more traditionally used for AI development.
Apple’s visual model is blazing fast
Now, Apple’s Machine Learning Research team has published FastVLM: a Visual Language Model (VLM) that leverages MLX to deliver nearly instantaneous high-resolution image processing, requiring significantly less computational power compared to similar models.
As Apple explains in its report:
Based on a comprehensive efficiency analysis of the interplay between image resolution, vision latency, token count, and LLM size, we introduce FastVLM—a model that achieves an optimized trade-off between latency, model size, and accuracy.
At the heart of FastVLM is an encoder named FastViTHD, designed specifically for efficient VLM performance on high-resolution images.
It's up to 3.2 times faster and 3.6 times smaller than comparable models. This is a significant advantage when aiming to process information directly on the device without relying on the cloud to generate a response to what the user has just asked or is looking at.
Moreover, FastVLM was designed to output fewer tokens, which is crucial during inference—the step where the model interprets the data and generates a response.
According to Apple, its model boasts an 85 times faster time-to-first-token compared to similar models, which is the time it takes for the user to input the first prompt and receive the first token of the answer. Fewer tokens on a faster and lighter model translate to swifter processing.
The FastVLM model is available on GitHub, and the report detailing its architecture and performance can be found on arXiv.
免责声明:info@kdj.com
所提供的信息并非交易建议。根据本文提供的信息进行的任何投资,kdj.com不承担任何责任。加密货币具有高波动性,强烈建议您深入研究后,谨慎投资!
如您认为本网站上使用的内容侵犯了您的版权,请立即联系我们(info@kdj.com),我们将及时删除。
-
-
-
-
- 加密鲸鱼堆积炒作:深入研究激进的积累和市场动态
- 2026-09-02 04:05:01
- 加密鲸鱼正在积极积累 HYPE 代币和狗狗币,这表明在散户抛售和机构持续兴趣的情况下,高信念趋势和潜在的市场转变。
-
- XRP 牛市分析师在市场周期中识别出“发射台”信号
- 2026-09-02 04:05:01
- XRP 牛市分析师 Thea Grace X 指出了一个关键的“发射台”水平,提供了对潜在市场周期和牛市陷阱的见解。
-
- 泰国 SEC 着眼于监管零售加密衍生品:新时代的曙光
- 2026-09-02 04:05:01
- 泰国 SEC 提出了一个新框架,允许散户投资者在监管下获取海外加密货币衍生品,这标志着加密货币监管的战略转变。
-
-
-

































