本文へ移動
cccskills
無料GitHub で公開

depth-estimation

Real-time depth map privacy transforms using Depth Anything v2 (CoreML + PyTorch)

インストール方法を見る

含まれるファイル(14)

  • SKILL.md3.7 KB
  • config.yaml2.6 KB
  • deploy.bat5.7 KB
  • deploy.sh7.3 KB
  • models.json4.9 KB
  • README.md3.0 KB
  • requirements_cpu.txt437 B
  • requirements_cuda.txt522 B
  • requirements_directml.txt518 B
  • requirements.txt1.3 KB
  • scripts/benchmark_coreml.py5.7 KB
  • scripts/benchmark.py11.3 KB
  • scripts/transform_base.py16.8 KB
  • scripts/transform.py25.9 KB

SKILL.md(原文)

インストールする前に、エージェントに与えられる指示の中身を確認できます。

Depth Estimation (Privacy)

Real-time monocular depth estimation using Depth Anything v2. Transforms camera feeds with colorized depth maps — near objects appear warm, far objects appear cool.

When used for privacy mode, the depth_only blend mode fully anonymizes the scene while preserving spatial layout and activity, enabling security monitoring without revealing identities.

Hardware Backends

PlatformBackendRuntimeModel
macOSCoreMLApple Neural Engineapple/coreml-depth-anything-v2-small (.mlpackage)
Linux/WindowsPyTorchCUDA / CPUdepth-anything/Depth-Anything-V2-Small (.pth)

On macOS, CoreML runs on the Neural Engine, leaving the GPU free for other tasks. The model is auto-downloaded from HuggingFace and stored at ~/.aegis-ai/models/feature-extraction/.

What You Get

  • Privacy anonymization — depth-only mode hides all visual identity
  • Depth overlays on live camera feeds
  • 3D scene understanding — spatial layout of the scene
  • CoreML acceleration — Neural Engine on Apple Silicon (3-5x faster than MPS)

Interface: TransformSkillBase

This skill implements the TransformSkillBase interface. Any new privacy skill can be created by subclassing TransformSkillBase and implementing two methods:

from transform_base import TransformSkillBase

class MyPrivacySkill(TransformSkillBase):
    def load_model(self, config):
        # Load your model, return {"model": "...", "device": "..."}
        ...

    def transform_frame(self, image, metadata):
        # Transform BGR image, return BGR image
        ...

Protocol

Aegis → Skill (stdin)

{"event": "frame", "frame_id": "cam1_1710001", "camera_id": "front_door", "frame_path": "/tmp/frame.jpg", "timestamp": "..."}
{"command": "config-update", "config": {"opacity": 0.8, "blend_mode": "overlay"}}
{"command": "stop"}

Skill → Aegis (stdout)

{"event": "ready", "model": "coreml-DepthAnythingV2SmallF16", "device": "neural_engine", "backend": "coreml"}
{"event": "transform", "frame_id": "cam1_1710001", "camera_id": "front_door", "transform_data": "<base64 JPEG>"}
{"event": "perf_stats", "total_frames": 50, "timings_ms": {"transform": {"avg": 12.5, ...}}}

Setup

python3 -m venv .venv && source .venv/bin/activate
pip install -r requirements.txt

レビュー

まだレビューはありません。使ってみた感想をお寄せください。

同じリポジトリのスキル

概要と使いどころ

Dataset annotation management — COCO labels, sequences, export, and Kaggle upload

日本語の概要は準備中です。原文の説明を表示しています。

SharpAI/DeepCamera3,0932026年9月17日 更新

Eufy camera integration — local RTSP streaming and event clips

日本語の概要は準備中です。原文の説明を表示しています。

SharpAI/DeepCamera3,0932026年9月17日 更新

Reolink camera integration — RTSP and HTTP API

日本語の概要は準備中です。原文の説明を表示しています。

SharpAI/DeepCamera3,0932026年9月17日 更新

TP-Link Tapo camera integration — RTSP streaming and ONVIF

日本語の概要は準備中です。原文の説明を表示しています。

SharpAI/DeepCamera3,0932026年9月17日 更新

LINE messaging channel for Clawdbot agent

日本語の概要は準備中です。原文の説明を表示しています。

SharpAI/DeepCamera3,0932026年9月17日 更新

Matrix/Element messaging channel for Clawdbot agent

日本語の概要は準備中です。原文の説明を表示しています。

SharpAI/DeepCamera3,0932026年9月17日 更新

SharpAI のスキルをすべて見る

このスキルの問題を報告する