Kaiju Hub AI

Agentic Agents

Agentic Agents area showing toolbar with add tab, markdown viewer, agents & skills, context history icons, Internal/External toggle, and empty state with Create Tab button

Overview

The Agentic Agents section is the core AI workspace within your project. It allows you to create and manage agent tabs — whether they are CLI-based, SDK-based, or ACP-based. Each tab runs an independent AI agent session, so you can work with multiple agents or conversations simultaneously.

Kaiju Hub AI ships with built-in support for Anthropic Claude Code, Google Gemini, and OpenAI Codex. The Pro version also supports custom CLI additions like OpenCode.

The agent that opens by default is determined by the provider and mode you selected in the Project Info area. Every time you left-click the + button, a new tab opens using that default configuration.

Kaiju Hub AI Context System

How It Works

Kaiju Hub AI uses a unique context bundling system to communicate with AI agents. Instead of sending raw messages directly, every prompt you send is bundled into a structured context file that the AI reads as a single package. This bundle includes your prompt text, any attached images, attached files, and references to blocks or projects you have included.

When you send a message, Kaiju Hub AI assembles all of these pieces into one bundle file and passes it to the AI agent. The agent reads the entire bundle at once, giving it full access to everything you have attached in a single, organized context.

What Gets Bundled

Prompt Text

Your message to the AI agent is included as the primary instruction.

Attached Images

Any images you attach are included in the bundle. You are not restricted to a small number — you can send many images in a single message.

Attached Files

Files you attach are bundled so the AI can read and reference their contents directly.

Project References

If you attach a project from the Project Hub, its path is included so the AI can reference and work with that project directly.

Block References

Blocks from the Block Library can be attached to a message. The AI will read and utilize the block content as part of its context.

Cross-Session Memory

Every bundle is stored on the backend after it is sent. This means the AI can reference context from previous tabs within the same session, and it can also look back at previous chat sessions entirely. For example, if you had a conversation last week about a specific feature, the AI can recall what you discussed and what you asked for.

This cross-session memory works across all tabs. If you send something in one tab, the context is available for the AI to reference from any other tab or future session.

Why This Matters

Reduced Token Usage

By bundling context into a structured file rather than appending raw history, the system cuts down on token consumption significantly.

More Context Material

You can include far more context — many images, multiple files, project references, and blocks — without hitting typical limitations.

Persistent History

The AI retains awareness of past conversations across sessions, so you do not need to repeat yourself or re-explain prior work.

Add Tab

Add Tab dropdown menu showing CLI, SDK, and ACP agent type options

Overview

The + button in the agent toolbar lets you create new agent tabs. Left-clicking the + button instantly opens a new tab using the default agent selected in the Project Info area. Right-clicking the + button opens a dropdown where you can choose a specific agent type.

Supported Agent Types

Claude CodeAnthropic
GeminiGoogle
CodexOpenAI
ACP AgentAgent Communication Protocol
Custom AgentPro only. Supports Internal and External modes.

Markdown Viewer

Markdown Viewer showing file selector with Edit, Split, and Preview modes, and empty state prompting to create a CLAUDE.md file

Overview

The Markdown Viewer tab renders markdown content within the agent area. It is used to display formatted documentation, agent instructions, skill files, and other markdown-based content directly in the workspace without switching to an external editor.

Features

Rich Rendering

Supports full markdown syntax including headings, lists, code blocks with syntax highlighting, tables, links, and images.

Agent Documentation

View agent instruction files (CLAUDE.md, agents/*.md, skills/*/SKILL.md) rendered in a readable format.

Inline Display

Content renders inline within the agent tab area, so you can read documentation alongside your active agent sessions.

Agent Area

Agent Area showing Agents and Skills tabs with options to create .claude/, .gemini/, and .codex/ provider directories

Overview

The Agent Area is the main workspace for interacting with AI agents. It displays chat messages, renders code with syntax highlighting, shows file and component context, and provides controls for managing the conversation. Each agent tab has its own independent agent area.

Tab Components

Code Terminal Tab

For CLI-based agents (Claude Code, Gemini CLI, Codex CLI). Shows a terminal-like interface where the agent executes commands and displays output in real-time.

SDK Chat Tab

For SDK-based agents (Claude SDK, Gemini SDK, OpenAI SDK). Provides a conversational chat interface for interacting with the agent.

ACP Chat Tab

For ACP (Agent Communication Protocol) agents. A specialized chat interface for communicating with dynamic protocol-based agents.

Agents & Skills Tab

Browse and manage agent capabilities, custom agents, and skill definitions for the current project.

Session History Panel

View a log of all past sessions for the current agent type, with the ability to load and review previous conversations.

Image Gallery

Browse images generated by AI agents during the current project session.

Context History

Context History panel showing a numbered list of past context entries with Prompt and Images tabs, timestamps, and image count badges

Overview

The Context History panel tracks all context that has been shared with AI agents throughout your project. It separates context into internal and external sources, giving you full visibility into what information agents have accessed.

Context Types

Internal Context

Context generated from within your project — selected files, checked code, component selections, and file content that you explicitly provided to agents during chat sessions.

External Context

Context from outside your project — web search results, fetched URLs, imported documentation, or other external data sources that agents referenced or that you attached.

Features

Chronological Log

Context entries are listed in chronological order so you can trace what was shared and when.

File Registry

An internal file registry tracks all files that have been shared with agents, including enhanced metadata and unique IDs for each file.

Context Indicators

The chat input area shows a context indicator displaying the number of files currently attached, so you always know what context is active.

Context Modal

Click the context indicator to open a detailed modal showing all current context — files, components, and projects — with options to add or remove items.

Internal & External

Agent toolbar showing External mode toggle with PowerShell shell selector dropdown

Overview

The Internal / External toggle in the agent toolbar controls how CLI-based agents are launched, including custom agents. This determines whether the agent runs inside the built-in workspace terminal or in a separate external terminal window on your system.

Modes

Internal

The agent runs inside the built-in workspace terminal. Output is displayed inline within the agent area, and the session is fully managed by Kaiju Hub AI.

External

The agent launches in a separate system terminal window (e.g. PowerShell, Terminal, iTerm). Useful when you need a full terminal environment or want to run agents alongside other command-line tools.

Shell Selector

When in External mode, a dropdown appears letting you choose which shell to use (e.g. PowerShell, Command Prompt, Git Bash). The selected shell is used when launching external CLI agent sessions.

Enhanced Chat Input

Enhanced Chat Input area showing text input field with sidebar icons for include context, project images, screen recordings, settings, and voice input, plus a send button

Overview

The Enhanced Chat Input is the main input area at the bottom of the agent workspace. It goes beyond a simple text box — offering image attachments, screen recording, file attachments, component selection, voice input, and granular permission controls for AI context.

Attachment Buttons

Gallery (Image Collection)

Opens the Image Upload Modal where you can browse, select, and upload images to include in your message. Supports multi-select with thumbnails and drag-and-drop upload.

Image Popup Modal

Click any attached image thumbnail to open it in a full-size centered overlay viewer. Close with the X button or by clicking outside the modal.

Screen Recording

Capture screen recordings directly from the chat input. The recorder captures frames at configurable FPS (30 or 60), tracks frame count and duration, and generates a manifest with metadata. Recorded frames can be attached to messages for visual context.

File Attachment

Attach project files directly to your message using the paperclip icon.

Workspace Context

Add workspace context using the brain icon — includes project files and metadata for richer AI understanding.

Component Selection

Select components or blocks from the library using the puzzle icon to include as context.

Include Context

Overview

The Include Context button is the first icon on the left sidebar of the chat input area. It allows you to attach project context — such as checked files, selected components, or workspace metadata — to your message before sending it to the AI agent.

How to Use

Left Click

Opens the context panel where you can browse and select which files and project context to include with your next message.

Right Click

Instantly includes the current context in your message without opening the panel. This is a quick shortcut to attach whatever files and context are currently selected.

Project Images

Project Images panel showing drag-and-drop upload area, image grid with thumbnails, selection controls (Select All, Deselect All, Delete All), and image count

Overview

The Project Images panel is a centralized gallery for all images associated with your project. Upload screenshots, mockups, reference images, or any visual assets you want to attach to AI conversations or keep organized within the workspace.

Features

Upload Images

Click the upload area or drag and drop images directly into the panel. You can also paste images from the clipboard using Ctrl+V / Cmd+V.

Image Grid

Uploaded images are displayed as a thumbnail grid for easy browsing and selection.

Selection Controls

Select individual images by clicking, or use Select All / Deselect All to manage bulk selections. Selected images can be deleted with Delete All.

Attach to Messages

Selected images can be included as visual context when sending messages to AI agents, giving them reference material for your requests.

Screen Recordings

Screen Recordings panel showing Capture Settings with Capture Speed (Slow 2 FPS / Normal 4 FPS), Frames Per Image slider, Grid Columns slider, and Capture Source selector with monitor details

Overview

The Screen Recordings panel lets you capture screen recordings and configure capture settings. Recorded frames are saved as images that can be attached to AI agent conversations, providing visual context for your requests.

Capture Settings

Capture Speed

Choose between Slow (2 FPS) and Normal (4 FPS). Use Slow if recording is too fast or you are missing content. Normal captures more detail but uses more storage.

Frames Per Image

Control how many frames are combined per image. Set to 1 for individual screenshots, or higher values for grid compositions that show multiple frames in a single image.

Grid Columns

Set the number of columns in the grid layout when combining multiple frames into a single image. Use 1 for vertical stacking or higher values for wide grids.

Capture Source

Select which screen or monitor to capture. Shows available monitors with resolution and position details. The primary monitor is labeled for easy identification.

Recordings Tab

The Recordings tab displays all previously captured recordings. Browse past sessions, review captured frames, and attach them to new AI conversations for visual reference.

Settings

Claude Code Settings panel showing Permission Mode with Skip All Permissions toggle and restart required notice

Overview

The Settings panel provides configuration options for the active AI agent. Access it from the gear icon on the chat input sidebar. Settings control permission modes and other agent-specific behavior.

Permission Mode

Skip All Permissions

When enabled, bypasses all permission checks for the agent (equivalent to --dangerously-skip-permissions). This lets the agent execute file edits, terminal commands, and other actions without confirmation prompts.

Restart Required

Changes to permission settings take effect on the next session. Close and reopen the agent tab to apply the new configuration.

Voice Input

Overview

The Voice Input button is the microphone icon at the bottom of the chat input sidebar. It provides platform-native speech-to-text functionality, letting you dictate messages to AI agents instead of typing.

How It Works

Activate

Click the microphone icon to start voice input. The icon shows an active state while listening.

Dictate

Speak naturally and your words are transcribed into the chat input field in real time using your system's speech recognition (Windows Speech on Windows, Dictation on macOS).

Send

Once done dictating, review the transcribed text and click send or press Enter to submit your message to the agent.

Shortcuts

Middle Mouse Button

On Windows and Mac (if your mouse has a middle button), press the middle mouse button to start recording. Click it again in the chat input area to stop recording and automatically send the message.

Custom Shortkey

You can set a custom keyboard shortcut for voice input in the Voice tab within Settings. This lets you start and stop voice recording without using the mouse.

All third-party logos, trademarks, and brand names displayed in Kaiju Hub AI are the property of their respective owners. Their use is solely for identification purposes and does not imply endorsement, sponsorship, or affiliation.