跳到正文
FunCoding

搜索

搜索文档、智能体、博客、Skill 和 MCP

cognee-integrations

Use when the user wants to connect cognee to external services — switching LLM or embedding providers (OpenAI, Azure, Gemini, Anthropic, Ollama, OpenRouter), changing databases (Postgres, PGVector, Neo4j, Neptune, Turso), S3 storage, or the MCP server for IDE integration.

数据库与数据32k.agents/skills/cognee-integrations/SKILL.md

安装

将以下指令发送给 Claude Code、Codex 或 Cursor,智能体会先检查内容的安全性,经你确认后再安装。

读取 https://funcoding.ai/skills/topoteretes/cognee/cognee-integrations/install.md ,按里面的步骤帮我安装这个 Skill。

SKILL.md

Set up cognee integrations

All integration config is environment variables (.env). The authoritative, always-current list with commented examples is .env.template at the repo root — check it before inventing variable names. Install the matching extra before switching a backend (e.g. pip install cognee[postgres]).

LLM providers

Default is OpenAI (LLM_API_KEY is all you need). To switch, set LLM_PROVIDER, LLM_MODEL, LLM_API_KEY, and (where relevant) LLM_ENDPOINT / LLM_API_VERSION:

  • Azure OpenAI: LLM_PROVIDER=azure, LLM_MODEL=azure/gpt-4o-mini, endpoint + api version required.
  • Gemini (no extra needed): LLM_PROVIDER=gemini, LLM_MODEL=gemini/gemini-2.0-flash-exp.
  • Anthropic (cognee[anthropic]): LLM_PROVIDER=anthropic, model e.g. claude-3-5-sonnet-20241022.
  • Ollama, local (cognee[ollama]): LLM_PROVIDER=ollama, LLM_ENDPOINT=http://localhost:11434/v1, and set the embedding block + HUGGINGFACE_TOKENIZER too.
  • Custom / OpenRouter / vLLM: LLM_PROVIDER=custom with the provider's OpenAI-compatible endpoint.
  • AWS Bedrock (cognee[aws]): LLM_PROVIDER=bedrock + AWS credentials/region.

The classic trap: LLM and embeddings are configured independently (EMBEDDING_PROVIDER, EMBEDDING_MODEL, EMBEDDING_ENDPOINT, EMBEDDING_API_KEY). Configuring only one leaves the other on OpenAI — either keep a valid OpenAI key or configure both.

Databases

  • Relational (DB_PROVIDER): sqlite (default) or postgres (cognee[postgres]; host/port/user/password/name via DB_* vars).
  • Vector (VECTOR_DB_PROVIDER): lancedb (default), pgvector (cognee[postgres], needs VECTOR_DB_URL), neptune_analytics (cognee[neptune]), turso (cognee[turso]). Anything else (ChromaDB, Qdrant, Weaviate, Milvus, …) lives in community adapters — install from https://github.com/topoteretes/cognee-community and register with use_vector_adapter before use; setting VECTOR_DB_PROVIDER alone raises "Unsupported vector database provider".
  • Graph (GRAPH_DATABASE_PROVIDER): ladybug (default), neo4j (cognee[neo4j], bolt URL + credentials), neptune (cognee[neptune]), ladybug-remote, postgres (no raw Cypher / natural-language search).

The repo docker-compose.yml ships ready-to-use postgres (pgvector) and neo4j profiles with matching default credentials. From a container, reach host services with DB_HOST=host.docker.internal.

Storage, cache, and the rest

  • S3 storage (cognee[aws]): STORAGE_BACKEND=s3 + bucket/credentials, and point DATA_ROOT_DIRECTORY/SYSTEM_ROOT_DIRECTORY at s3:// paths.
  • Session cache: CACHE_BACKEND = sqlite (default) | postgres | redis | fs | tapes.
  • Ontologies: ONTOLOGY_FILE_PATH to an OWL file, resolver/matching via ONTOLOGY_RESOLVER / MATCHING_STRATEGY.

MCP server (IDE integration)

docker compose --profile mcp up starts the MCP server on port 8001 (Streamable HTTP at http://localhost:8001/mcp), built from cognee-mcp/. Point Cursor / Claude Desktop / Claude Code at it to use cognee memory from the IDE. Configure its DB_* env to match the main service so both see the same data.

After changing providers mid-project

Embeddings from different models are not comparable — after switching the embedding provider or model, reset local state (cognee-cli forget --all or await cognee.forget(everything=True)) and re-ingest with remember().

To drop just the graph and vectors while keeping the ingested files, use await cognee.forget(dataset="my_project", memory_only=True) — the dataset can then be rebuilt under the new embedding model without re-uploading anything.

相似的 Skill

xlsx
官方
anthropics/skills180k

xlsx

Use this skill any time a spreadsheet file is the primary input or output. This means any task where the user wants to: open, read, edit, or fix an existing .xlsx, .xlsm, .xltx, .csv, or .tsv file (e.g., adding columns, computing formulas, formatting, charting, cleaning messy data); create a new spreadsheet from scratch or from other data sources; or convert between tabular file formats. Trigger especially when the user references a spreadsheet file by name or path — even casually (like "the xlsx in my downloads") — and wants something done to it or produced from it. Also trigger for cleaning or restructuring messy tabular data files (malformed rows, misplaced headers, junk data) into proper spreadsheets. The deliverable must be a spreadsheet file. Do NOT trigger when the primary deliverable is a Word document, HTML report, standalone Python script, database pipeline, or Google Sheets API integration, even if tabular data is involved.

数据库与数据

deprecation-and-migration
addyosmani/agent-skills102k

deprecation-and-migration

Manages deprecation and migration. Use when removing old systems, APIs, or features. Use when migrating users from one implementation to another. Use when migrating a database schema in production, such as renaming or dropping a column without downtime (expand/contract). Use when deciding whether to maintain or sunset existing code.

数据库与数据

smart-explore
thedotmack/claude-mem97k

smart-explore

Token-optimized structural code search using tree-sitter AST parsing. Use instead of reading full files when you need to understand code structure, find functions, or explore a codebase efficiently.

数据库与数据

pathfinder
thedotmack/claude-mem97k

pathfinder

Map a codebase into feature-grouped flowcharts, identify duplicated concerns across features, and propose a unified architecture. Use when asked to "find the ideal path," unify duplicated systems, or audit architecture before a refactor. Emits a proposed unified flowchart plus per-system /make-plan prompts.

数据库与数据

oh-my-issues
thedotmack/claude-mem97k

oh-my-issues

Cluster a GitHub issue backlog by root cause into a small set of plan-master issues, redirect children with a standardized comment, and bundle architectural-fix PRs that close clusters atomically. Use when an issue tracker has accumulated dozens of reports that share underlying defects, when asked to triage / consolidate / cluster / dedupe issues, when asked to build a plan series or roadmap from open issues, or when routing a new incoming bug into an existing plan.

数据库与数据

mode-creator
thedotmack/claude-mem97k

mode-creator

Interactively create, install, activate, and verify custom claude-mem modes, including domain-specific observation types, concept tags, optional Telegram alerts, bot setup, worker restart, and startup-context verification. Use this whenever someone asks to customize what claude-mem remembers, create or change a mode, track domain-specific notes, add observation types or tags, or send Telegram notifications for particular memories—even if they do not use the word "mode."

数据库与数据