時価総額: $2.6355T 0.69%
ボリューム(24時間): $41.7705B -52.14%
  • 時価総額: $2.6355T 0.69%
  • ボリューム(24時間): $41.7705B -52.14%
  • 恐怖と貪欲の指数:
  • 時価総額: $2.6355T 0.69%
暗号
トピック
暗号化
ニュース
暗号造園
動画
トップニュース
暗号
トピック
暗号化
ニュース
暗号造園
動画
bitcoin
bitcoin

$78255.876969 USD

1.01%

ethereum
ethereum

$2459.268476 USD

1.00%

tether
tether

$0.999934 USD

-0.01%

bnb
bnb

$693.927389 USD

0.93%

xrp
xrp

$1.400219 USD

1.31%

usd-coin
usd-coin

$0.999998 USD

0.00%

solana
solana

$105.298299 USD

1.53%

tron
tron

$0.340578 USD

0.66%

hyperliquid
hyperliquid

$83.234542 USD

2.36%

dogecoin
dogecoin

$0.084950 USD

0.56%

zcash
zcash

$836.980796 USD

5.15%

unus-sed-leo
unus-sed-leo

$9.709987 USD

0.25%

monero
monero

$474.403454 USD

1.57%

chainlink
chainlink

$11.436884 USD

0.87%

cardano
cardano

$0.201963 USD

0.79%

暗号通貨のニュース記事

Apple's new visual model is FAST

2025/05/13 05:59

Apple's new visual model is FAST

Apple has been busy developing its own AI technologies, and recently offered a glimpse into how its models might work.

Currently, Apple’s direct competitors to the Meta Ray-Bans are planned for around 2027, together with AirPods equipped with cameras, which will provide their own set of AI-enabled capabilities.

While it’s still too early to anticipate what they will precisely look like, Apple unveiled MLX, its own open ML framework designed specifically for Apple Silicon.

Essentially, MLX provides a lightweight method to train and run models directly on Apple devices, remaining familiar to developers who prefer frameworks and languages more traditionally used for AI development.

Apple’s visual model is blazing fast

Now, Apple’s Machine Learning Research team has published FastVLM: a Visual Language Model (VLM) that leverages MLX to deliver nearly instantaneous high-resolution image processing, requiring significantly less computational power compared to similar models.

As Apple explains in its report:

Based on a comprehensive efficiency analysis of the interplay between image resolution, vision latency, token count, and LLM size, we introduce FastVLM—a model that achieves an optimized trade-off between latency, model size, and accuracy.

At the heart of FastVLM is an encoder named FastViTHD, designed specifically for efficient VLM performance on high-resolution images.

It's up to 3.2 times faster and 3.6 times smaller than comparable models. This is a significant advantage when aiming to process information directly on the device without relying on the cloud to generate a response to what the user has just asked or is looking at.

Moreover, FastVLM was designed to output fewer tokens, which is crucial during inference—the step where the model interprets the data and generates a response.

According to Apple, its model boasts an 85 times faster time-to-first-token compared to similar models, which is the time it takes for the user to input the first prompt and receive the first token of the answer. Fewer tokens on a faster and lighter model translate to swifter processing.

The FastVLM model is available on GitHub, and the report detailing its architecture and performance can be found on arXiv.

オリジナルソース:9to5mac

免責事項:info@kdj.com

提供される情報は取引に関するアドバイスではありません。 kdj.com は、この記事で提供される情報に基づいて行われた投資に対して一切の責任を負いません。暗号通貨は変動性が高いため、十分な調査を行った上で慎重に投資することを強くお勧めします。

このウェブサイトで使用されているコンテンツが著作権を侵害していると思われる場合は、直ちに当社 (info@kdj.com) までご連絡ください。速やかに削除させていただきます。

2026年08月31日 に掲載されたその他の記事