How to Use DeepSeek V4 Flash in Glarity: Translate Pages, Summarize Pages & YouTube Videos

Connect your own DeepSeek V4 Flash API key to Glarity in five minutes: the OpenAI API key option, custom model and endpoint fields you need to fill, and everything it unlocks.

You can run Glarity on DeepSeek's V4 Flash model in five minutes: open Settings → General, choose OpenAI API key (it accepts custom keys from any AI vendor), paste your DeepSeek key, and fill in the model, context window, API host, path, and temperature. Once connected, page translation, page summaries, and YouTube video summaries all run through your own DeepSeek endpoint.

Why connect your own DeepSeek V4 Flash key

Glarity's built-in "Glarity (Recommended for All Users)" option gives you 20,000 free tokens per day. That is plenty for casual use, but a custom API key removes the daily cap: your DeepSeek account bills the usage directly, and Glarity's summaries and translations keep running as long as your key does.

DeepSeek V4 Flash is a fast, low-cost model with a large context window, which makes it a good fit for the long inputs Glarity handles — full web pages, articles, and hour-long video transcripts. The exact model ID is deepseek-v4-flash.

What you need before you start

  1. A DeepSeek account and API key from the official console at platform.deepseek.com (official DeepSeek platform — API keys are created there; the model is billed by DeepSeek on usage).
  2. The four connection values shown in this guide (model, context window, host, path) — every one of them is visible in the screenshot below.

Configure Glarity step by step

Open the Glarity settings window, go to General, and scroll to Connect to AI for quick Glarity summaries and translations. This page is where every AI service key lives.

Glarity settings connected to a DeepSeek API key: custom model, 204,800-token context window, api.deepseek.com host, /v1/chat/completions path, temperature 0.5
  1. Pick "OpenAI API key" — the third option on the page, labeled "Support for custom keys from any AI vendor". This is the one that accepts DeepSeek.
  2. Paste your DeepSeek API key into the API Key field.
  3. Model — choose Custom Model from the dropdown and type deepseek-v4-flash. (If you later want a different model, this is where you change it — a DeepSeek key also works with deepseek-v4-pro.)
  4. Model Context Window — select 204800 tokens. This is the window the extension tells the model to use; it comfortably covers full-page summaries and long video transcripts. (DeepSeek's own API supports up to 1M tokens, so you can raise this later if Glarity's menu offers it.)
  5. API Host — enter https://api.deepseek.com (DeepSeek's official OpenAI-compatible endpoint).
  6. API Path — enter /v1/chat/completions. (DeepSeek's docs sample the same OpenAI-compatible endpoint as /chat/completions; both work because DeepSeek follows OpenAI's API format — the values in the screenshot are the tested combination.)
  7. Temperature — set 0.5. This default keeps summaries grounded and reproducible; raise it if you want more varied phrasing.
  8. Click Test to verify the connection, then Save to activate it for every Glarity summary and translation feature.

The settings at a glance

Field

Value

Why

Connection

OpenAI API key

Custom keys from any vendor, OpenAI-compatible

Model

Custom Model — deepseek-v4-flash

DeepSeek's fast low-cost model ID

Model Context Window

204800 tokens

Fits long pages and full video transcripts

API Host

https://api.deepseek.com

Official DeepSeek API base URL

API Path

/v1/chat/completions

OpenAI-compatible chat completion route

Temperature

0.5

Balanced, grounded output

What this unlocks

With the key saved, every Glarity capability routes through DeepSeek V4 Flash:

Capability

What it does now

Translate

Immersive page translation and side-by-side subtitles — the same feature set covered in our bilingual subtitle tutorial

Page Summary

Summaries, key points, and follow-up questions for any web page

YouTube Summary

Transcript-based summaries with timestamped highlights, key points, and an FAQ for any video

Video Subtitles

Generate and translate video subtitles in your language

Quick Email Reply

Drafted replies in the tone and language you need

1-Click Prompts

Favorite tasks run again on your own model

Cost and context notes

DeepSeek bills usage; the official pricing page lists V4 Flash at round $0.22–0.44 per 1M input tokens (cache miss) and $0.66–1.32 per 1M output tokens, with off-peak and peak rates — check the official pricing page for current numbers. Typical Glarity requests (one page summary, one video transcript) are a small fraction of a million tokens, so day-to-day use is usually cents — but your balances and limits are under your DeepSeek account, not Glarity's.

Troubleshooting

  • Test fails with 401 / "invalid key" — the API key is wrong, expired, or copied with spaces. Recreate it on platform.deepseek.com and paste again.
  • Test fails with payment or balance errors — your DeepSeek account needs a top-up; the model is billed against your account.
  • "Model not found" — the model string is off. Use deepseek-v4-flash exactly; the docs note the model has been updated (V4-Flash-0731) but the caller string stays deepseek-v4-flash.
  • Slow or interrupted responses — a 429 rate limit or concurrency peak on DeepSeek's side; retry or wait, the connection itself is fine.
  • Nothing changed after saving — make sure you clicked Save after Test; the connection activates per settings window.

FAQ

Does DeepSeek also power translations, or only summaries? Both. The "Connect to AI" page covers Glarity's quick summaries and translations — Web page translation, selected text, and video subtitle translation all use the same connected model.

Is this free? You keep Glarity's free 20,000 tokens/day built-in option if you don't connect a key. Once you connect your DeepSeek key, usage is billed by DeepSeek (pay-as-you-go on your account) — see the official pricing page for rates.

Why a custom model instead of "Glarity" in the dropdown? The built-in Glarity option includes a free daily token allowance. The custom key route is for when you want your own provider, your own budget, no daily cap — or simply a model you already know.

Can I swap models later without a new key? Yes. The same key works with any DeepSeek model (for example deepseek-v4-pro), you just change the Model field, and if you connect a different OpenAI-compatible vendor later, you would update the host and path too.

Do I have to run the connection test? No, but it is the fast way to catch a typo in the host, path, or model ID before you start using summaries.

By the Glarity Editorial Team. The Glarity Editorial Team writes about AI search, video summarization, and getting more from your browser.

Glarity is a free browser extension for AI search and YouTube summarization — try it at glarity.app.