GLAMMBOX

The models, and the brain behind them.

A company makes models. An app gives you a place to talk to one. An agent gives a model tools and lets it work through several steps. A brain is separate again: it stores information the model can search later.

COMPANY → APPLICATION → MODEL → AGENT

Token

A small piece of text. Prices and context limits are usually counted in tokens, not pages.

Context window

How much the model can consider at once. A larger window does not guarantee a better answer.

Weights

The trained numbers inside a model. Open weights can be downloaded; closed weights stay on the provider’s servers.

Local or cloud

Local runs on your machine. Cloud runs elsewhere through an app or API. Local is not automatically private; every connection still matters.

API

A documented way for software to send a job to a model and receive the result.

Parameters

A model-size measure—not a quality score. Test models on your real job instead of choosing the biggest number.

Embedding

Turning a sentence into numbers so similar ideas sit near each other, even if they do not share the same words.

Qdrant · meaning shelves

A store of those numbers. Ask in your own words; it brings back notes that mean the same thing.

BM25 · exact words

A back-of-the-book index. It finds a name, a date, or a model number exactly as written.

Neo4j · relationship map

An optional map of who is tied to what. Useful for “who decided this?” — not required for ordinary search.

RAG

Retrieve first, then generate: the program finds notes, then the model writes the answer from those notes.

MCP · the plug

A plug that lets a model use a tool (files, mail, a calendar) instead of only talking about it. Credentials stay on the host.

Model information checked against official sources on 2 October 2026. Preview films illustrate the providers and may show earlier interfaces. Access, prices and release status can change; use the source links to verify your choice.

EVERYDAY CLOUD MODELS

Illustrative preview · current model details below

ANTHROPIC

Claude

Opus 5.5 · Sonnet 5.5 · Fable 5.1

Claude is Anthropic’s family for writing, analysis and coding. Compare the available model in your Claude plan or developer account against the work you actually need to do.

Access and official sources

Opus 5.5, Sonnet 5.5 and Fable 5.1 are released. Haiku 5.5 is announced for later; Haiku 4.5 remains the current released Haiku. Mythos 5.1 is restricted to vetted organizations. Availability and pricing depend on the product.

Checked 2 October 2026 · Official provider sources.

Official source 1 · Official source 2 · Official source 3

Illustrative preview · current model details below

OPENAI

GPT

GPT-6 Astra · GPT-6.1 Sol · GPT-6 Luna

OpenAI’s GPT family supports reasoning, writing and coding. ChatGPT is an application; Codex provides a coding workflow; the API lets developers build model capabilities into software.

Access and official sources

The official API catalog lists gpt-6-astra, gpt-6.1-sol and gpt-6-luna. Product access, model choices and billing differ between ChatGPT, Codex and the API. OpenAI also publishes a separate open-weight gpt-oss family.

Checked 2 October 2026 · Official provider sources.

Official source 1

Archive clip · Grok 4.5

XAI

Grok

Grok 4.7

Grok is xAI’s model family for reasoning and work with connected tools. An application may provide search or other tools; check the specific product rather than assuming every model sees live information.

Access and official sources

Grok 4.7 is available through the public API as grok-4.7. Grok 4.7 Fast is offered in Cursor and Grok Build, not the public API catalog at this check. Subscription and API billing are separate access routes.

Checked 2 October 2026 · Official provider sources.

Official source 1

Separate showcase · Gemini Robotics 2

GOOGLE

Gemini

Gemini 3.8 Flash · 3.1 Pro Preview

Google’s Gemini family handles tasks involving text and media. It is useful to compare the fast Flash route with Pro for the complexity of your documents, images or video.

Access and official sources

Gemini 3.8 Flash is stable in the API; Gemini 3.1 Pro is a preview. Gemini 4 Argon is announced with restricted access for trusted cyberdefense partners, not a general public replacement. Veo is a separate video-generation family.

Checked 2 October 2026 · Official provider sources.

Official source 1 · Official source 2

DOWNLOADABLE AND LOWER-COST OPTIONS

These families matter when you want downloadable weights, lower API cost, very long context, or media tools under the same provider. Test the exact version you plan to use.

Illustrative preview · current model details below

ALIBABA · QWEN

Qwen

Qwen 3.8-Max

Alibaba’s Qwen family offers hosted models and separately licensed downloadable models. Qwen 3.8-Max accepts text, images and video and produces text through the hosted API.

Access and official sources

The hosted API name is qwen3.8-max, with the updated qwen3.8-max-0902 snapshot. The original release was in August; September 2 was an update. Check the exact downloadable variant, license and hardware needs separately; hosted capability does not prove laptop suitability.

Checked 2 October 2026 · Official provider sources.

Official source 1 · Official source 2

Illustrative preview · current model details below

MOONSHOT AI

Kimi

Kimi K3

Moonshot AI’s Kimi K3 combines native vision with coding and knowledge work. Consider it for tasks that bring documents, visual material and implementation together.

Access and official sources

The official release identifies kimi-k3 for API access. Check product limits, hosting options and any model license for the exact version. Very large context and model-size numbers are not a substitute for testing your own documents.

Checked 2 October 2026 · Official provider sources.

Official source 1 · Official source 2

DEEPSEEK

DeepSeek

V4.1 Flash · V4 Pro

DeepSeek offers models for reasoning and coding through its application and developer API. V4.1 Flash adds native multimodal input; compare the current model against your task and budget.

Access and official sources

V4.1 Flash was released September 10 and uses deepseek-flash. V4 Pro remains available as deepseek-v4-pro. Older Flash names are retired or compatibility aliases. Consult current pricing instead of assuming it is the cheapest provider.

Checked 2 October 2026 · Official provider sources.

Official source 1

MINIMAX

MiniMax

MiniMax M3 · H3 · Music 3

MiniMax M3 combines image and video understanding with coding and agent tasks. MiniMax also offers separate tools for making video and music.

Access and official sources

MiniMax-M3 is the current M3 API identifier. H3 and Music 3 serve separate media-generation roles; M3 itself should not be presented as a music generator. Check availability, usage costs and licensing for each product.

Checked 2 October 2026 · Official provider sources.

Official source 1

ZHIPU · Z.AI

GLM

GLM 5.3 · GLM 5.3 Flash

Z.ai’s GLM family supports coding and agent workflows. It offers a hosted route and downloadable weights for teams that want to manage their own deployment.

Access and official sources

The hosted model is glm-5.3. GLM 5.3 weights are already published by the official organization; GLM 5.3 Flash is a separate variant. Check the model card and license before planning a local installation.

Checked 2 October 2026 · Official provider sources.

Official source 1 · Official source 2

MODELS BUILT INTO PLATFORMS AND HARDWARE

Illustrative preview · current model details below

META

Meta AI

Muse Spark 1.3 · Muse Code · Glimmer

Meta’s model lineup includes Muse Spark, a separate Muse Code product and the open-weight Glimmer family. Match the product to the work rather than treating every Meta model as the same service.

Access and official sources

The official model index identifies Muse Spark 1.3 and these specialist families. Access varies by product. The index is verified; detailed availability and pricing should be checked on the linked provider page before committing.

Checked 2 October 2026 · Official provider sources.

Official source 1

Illustrative preview · current model details below

NVIDIA

Nemotron

Nemotron 3.5 Lightning 30B A3B

NVIDIA’s Nemotron models are building blocks for developers who want to deploy AI services. Hosted NIM access and self-hosting offer different ways to run the model.

Access and official sources

The verified model card is nvidia/nemotron-3.5-lightning-30b-a3b. Open weights and published recipes do not mean all training data is public. Hosting still needs suitable hardware, configuration and operation.

Checked 2 October 2026 · Official provider sources.

Official source 1

Illustrative preview · current model details below

MICROSOFT AI

MAI

Thinking-1 · Code-1.1 Flash · Image-2.6 · Voice-2.1 · Transcribe-2

Microsoft AI offers specialist models for reasoning, code, images, speech generation and transcription. Choose the model for the task: transcribing a recording and generating a voice are different jobs.

Access and official sources

The current official catalog lists these specialist families, including Voice 2.1 and Image 2.6. Product availability and plans vary. Do not assume every model is already installed in Copilot or VS Code, or included with an existing subscription.

Checked 2 October 2026 · Official provider sources.

Official source 1 · Official source 2

1. Choose your files

2. Find a useful passage

3. Keep it for your next project

A model answers. Your memory keeps the source.

Learn · GLAMMBRAIN

Your AI can change. Your memory stays yours.

Keep the files, decisions and project notes you choose on your own machine. Search them and give useful passages to an AI tool you explicitly connect. The memory starts empty; you control what enters it.

A memory starts with sources you choose. Retrieved passages can help a connected AI answer, while you control what is shared and what needs your approval.

Give this to your bot

Help me organize selected notes and documents so I can find useful information and its sources again. Explain my options, what stays private and what needs my approval.

Your files. Your context.

A Brain is a searchable memory built from sources you choose: notes, decisions and project records. It helps you find useful context again when a conversation ends or you change AI tools.

The source and the index are different.

Your original files hold the information. A search index is a derived guide that helps locate relevant passages by meaning or exact words. It is not a backup. Changes to a source may not appear until the index is refreshed.

Find the evidence. Then answer.

A connected AI can use retrieved passages to draft an answer and point back to the source. Check the source, its date and its context. Missing or conflicting evidence should remain visible; retrieval does not guarantee a correct answer.

You decide what is shared.

Reading this page does not give an AI access to your files. Access depends on the tool, its connections and your permission. A local source can still be sent to an online provider if you share it there. Review what leaves your device and who may receive it.

Options first. Your approval next.

An agent should explain suitable options and their limits, then ask before installing software, uploading documents or changing security settings. Its available tools vary. A Brain does not automatically connect every AI or remember every conversation.

SYSTEM

Each point is a file, a fact, a concept or a department; each line is a relationship recorded between two of them.

    HOW A RELATIONSHIP MAP CONNECTS PEOPLE, PROJECTS, AND DECISIONSGraph

    Swipe or scroll to keep reading. Use the arrows and zoom buttons to explore; on a computer, you can also drag with the mouse.

    This illustrative graph uses generic labels. It shows relationships, not private records or automatic access.