時価総額: $2.2026T 0.80%
ボリューム(24時間): $38.3583B -35.30%
  • 時価総額: $2.2026T 0.80%
  • ボリューム(24時間): $38.3583B -35.30%
  • 恐怖と貪欲の指数:
  • 時価総額: $2.2026T 0.80%
暗号
トピック
暗号化
ニュース
暗号造園
動画
トップニュース
暗号
トピック
暗号化
ニュース
暗号造園
動画
bitcoin
bitcoin

$87959.907984 USD

1.34%

ethereum
ethereum

$2920.497338 USD

3.04%

tether
tether

$0.999775 USD

0.00%

xrp
xrp

$2.237324 USD

8.12%

bnb
bnb

$860.243768 USD

0.90%

solana
solana

$138.089498 USD

5.43%

usd-coin
usd-coin

$0.999807 USD

0.01%

tron
tron

$0.272801 USD

-1.53%

dogecoin
dogecoin

$0.150904 USD

2.96%

cardano
cardano

$0.421635 USD

1.97%

hyperliquid
hyperliquid

$32.152445 USD

2.23%

bitcoin-cash
bitcoin-cash

$533.301069 USD

-1.94%

chainlink
chainlink

$12.953417 USD

2.68%

unus-sed-leo
unus-sed-leo

$9.535951 USD

0.73%

zcash
zcash

$521.483386 USD

-2.87%

暗号通貨のニュース記事

NVIDIA は、Apple によるより高速で優れた AI エクスペリエンスの構築を支援しています

2024/12/20 19:52

BGR リンクを通じて購入すると、当社はアフィリエイト手数料を獲得し、専門家の製品ラボのサポートに役立てることができます。 Apple と NVIDIA がコラボレーションの詳細を共有

NVIDIA は、Apple によるより高速で優れた AI エクスペリエンスの構築を支援しています

Tech giants Apple and NVIDIA have joined forces to enhance the performance of Large Language Models (LLMs) by introducing a new text generation technique for AI.

テクノロジー大手の Apple と NVIDIA は、AI 用の新しいテキスト生成技術を導入することで大規模言語モデル (LLM) のパフォーマンスを強化するために提携しました。

According to Apple, accelerating LLM inference is a crucial ML research problem. This is because auto-regressive token generation is computationally expensive and relatively slow. As a result, improving inference efficiency can reduce latency for users.

Apple によれば、LLM 推論の高速化は ML 研究の重要な課題です。これは、自動回帰トークンの生成は計算コストが高く、比較的時間がかかるためです。結果として、推論効率を向上させることで、ユーザーの待ち時間を短縮できます。

In addition to ongoing efforts to accelerate inference on Apple silicon, the company has recently made significant progress in accelerating LLM inference for the NVIDIA GPUs widely used for production applications across the industry, the company writes in a research paper.

Apple シリコンでの推論を高速化する継続的な取り組みに加え、同社は最近、業界全体の実稼働アプリケーションに広く使用されている NVIDIA GPU の LLM 推論の高速化において大きな進歩を遂げたと研究論文で述べています。

Earlier this year, Apple published and open-sourced Recurrent Drafter (ReDrafter), which is a novel approach to speculative decoding that “achieves state of the art performance.”

今年の初め、Apple は Recurrent Drafter (ReDrafter) を公開し、オープンソース化しました。これは、「最先端のパフォーマンスを実現する」投機的デコーディングへの新しいアプローチです。

According to the company, ReDrafter uses an RNN draft model, and combines beam search with dynamic tree attention to speed up LLM token generation by up to 3.5 tokens per generation step for open source models, surpassing the performance of prior speculative decoding techniques.

同社によれば、ReDrafter は RNN ドラフト モデルを使用し、ビーム検索と動的ツリー アテンションを組み合わせて、オープンソース モデルの LLM トークン生成を生成ステップあたり最大 3.5 トークンまで高速化し、以前の投機的デコード技術のパフォーマンスを上回りました。

“In benchmarking a tens-of-billions parameter production model on NVIDIA GPUs, using the NVIDIA TensorRT-LLM inference acceleration framework with ReDrafter, we have seen 2.7x speed-up in generated tokens per second for greedy decoding,” Apple papers show.

Apple の論文には、「NVIDIA TensorRT-LLM 推論アクセラレーション フレームワークと ReDrafter を使用して、NVIDIA GPU で数百億のパラメータ生成モデルのベンチマークを行ったところ、貪欲なデコードで 1 秒あたり生成されるトークンの速度が 2.7 倍向上したことがわかりました」と記載されています。

With that, this technology could signifanctly reduce latency users may experience, while also using fewer GPUs and consuming less power.

これにより、このテクノロジーはユーザーが経験する可能性のある遅延を大幅に削減すると同時に、使用する GPU と消費電力を削減することができます。

オリジナルソース:bgr

免責事項:info@kdj.com

提供される情報は取引に関するアドバイスではありません。 kdj.com は、この記事で提供される情報に基づいて行われた投資に対して一切の責任を負いません。暗号通貨は変動性が高いため、十分な調査を行った上で慎重に投資することを強くお勧めします。

このウェブサイトで使用されているコンテンツが著作権を侵害していると思われる場合は、直ちに当社 (info@kdj.com) までご連絡ください。速やかに削除させていただきます。

2026年07月26日 に掲載されたその他の記事