← 返回 Skills 市场
whitejohnk-26

Italiano Photo Video Maker

作者 whitejohnk-26 · GitHub ↗ · v1.0.0 · MIT-0
cross-platform ⚠ suspicious
71
总下载
0
收藏
0
当前安装
1
版本数
在 OpenClaw 中安装
/install italiano-photo-video-maker
功能描述
create photos and images into photo slideshow video with this italiano-photo-video-maker skill. Works with JPG, PNG, HEIC, MP4 files up to 200MB. Italian con...
使用说明 (SKILL.md)

Getting Started

Share your photos and images and I'll get started on AI video creation. Or just tell me what you're thinking.

Try saying:

  • "create my photos and images"
  • "export 1080p MP4"
  • "create a slideshow video with Italian"

Getting Connected

Before handling any user request, establish a connection to the backend API. Show a brief status like "Connecting...".

If NEMO_TOKEN is in the environment, use it directly and create a session. Otherwise, acquire a free starter token:

  • Generate a UUID as client identifier
  • POST to https://mega-api-prod.nemovideo.ai/api/auth/anonymous-token with the X-Client-Id header
  • The response includes a token with 100 free credits valid for 7 days — use it as NEMO_TOKEN

Then create a session by POSTing to https://mega-api-prod.nemovideo.ai/api/tasks/me/with-session/nemo_agent with Bearer authorization and body {"task_name":"project","language":"en"}. The session_id in the response is needed for all following requests.

Tell the user you're ready. Keep the technical details out of the chat.

Italiano Photo Video Maker — Turn Photos Into Videos

Drop your photos and images in the chat and tell me what you need. I'll handle the AI video creation on cloud GPUs — you don't need anything installed locally.

Here's a typical use: you send a ten vacation photos from Italy, ask for create a slideshow video with Italian music and transitions, and about 30-60 seconds later you've got a MP4 file ready to download. The whole thing runs at 1080p by default.

One thing worth knowing — fewer photos per video means faster processing and smoother transitions.

Matching Input to Actions

User prompts referencing italiano photo video maker, aspect ratio, text overlays, or audio tracks get routed to the corresponding action via keyword and intent classification.

User says... Action Skip SSE?
"export" / "导出" / "download" / "send me the video" → §3.5 Export
"credits" / "积分" / "balance" / "余额" → §3.3 Credits
"status" / "状态" / "show tracks" → §3.4 State
"upload" / "上传" / user sends file → §3.2 Upload
Everything else (generate, edit, add BGM…) → §3.1 SSE

Cloud Render Pipeline Details

Each export job queues on a cloud GPU node that composites video layers, applies platform-spec compression (H.264, up to 1080x1920), and returns a download URL within 30-90 seconds. The session token carries render job IDs, so closing the tab before completion orphans the job.

All calls go to https://mega-api-prod.nemovideo.ai. The main endpoints:

  1. SessionPOST /api/tasks/me/with-session/nemo_agent with {"task_name":"project","language":"\x3Clang>"}. Gives you a session_id.
  2. Chat (SSE)POST /run_sse with session_id and your message in new_message.parts[0].text. Set Accept: text/event-stream. Up to 15 min.
  3. UploadPOST /api/upload-video/nemo_agent/me/\x3Csid> — multipart file or JSON with URLs.
  4. CreditsGET /api/credits/balance/simple — returns available, frozen, total.
  5. StateGET /api/state/nemo_agent/me/\x3Csid>/latest — current draft and media info.
  6. ExportPOST /api/render/proxy/lambda with render ID and draft JSON. Poll GET /api/render/proxy/lambda/\x3Cid> every 30s for completed status and download URL.

Formats: mp4, mov, avi, webm, mkv, jpg, png, gif, webp, mp3, wav, m4a, aac.

Skill attribution — read from this file's YAML frontmatter at runtime:

  • X-Skill-Source: italiano-photo-video-maker
  • X-Skill-Version: from frontmatter version
  • X-Skill-Platform: detect from install path (~/.clawhub/clawhub, ~/.cursor/skills/cursor, else unknown)

Every API call needs Authorization: Bearer \x3CNEMO_TOKEN> plus the three attribution headers above. If any header is missing, exports return 402.

Draft field mapping: t=tracks, tt=track type (0=video, 1=audio, 7=text), sg=segments, d=duration(ms), m=metadata.

Timeline (3 tracks): 1. Video: city timelapse (0-10s) 2. BGM: Lo-fi (0-10s, 35%) 3. Title: "Urban Dreams" (0-3s)

Backend Response Translation

The backend assumes a GUI exists. Translate these into API actions:

Backend says You do
"click [button]" / "点击" Execute via API
"open [panel]" / "打开" Query session state
"drag/drop" / "拖拽" Send edit via SSE
"preview in timeline" Show track summary
"Export button" / "导出" Execute export workflow

SSE Event Handling

Event Action
Text response Apply GUI translation (§4), present to user
Tool call/result Process internally, don't forward
heartbeat / empty data: Keep waiting. Every 2 min: "⏳ Still working..."
Stream closes Process final response

~30% of editing operations return no text in the SSE stream. When this happens: poll session state to verify the edit was applied, then summarize changes to the user.

Error Handling

Code Meaning Action
0 Success Continue
1001 Bad/expired token Re-auth via anonymous-token (tokens expire after 7 days)
1002 Session not found New session §3.0
2001 No credits Anonymous: show registration URL with ?bind=\x3Cid> (get \x3Cid> from create-session or state response when needed). Registered: "Top up credits in your account"
4001 Unsupported file Show supported formats
4002 File too large Suggest compress/trim
400 Missing X-Client-Id Generate Client-Id and retry (see §1)
402 Free plan export blocked Subscription tier issue, NOT credits. "Register or upgrade your plan to unlock export."
429 Rate limit (1 token/client/7 days) Retry in 30s once

Common Workflows

Quick edit: Upload → "create a slideshow video with Italian music and transitions" → Download MP4. Takes 30-60 seconds for a 30-second clip.

Batch style: Upload multiple files in one session. Process them one by one with different instructions. Each gets its own render.

Iterative: Start with a rough cut, preview the result, then refine. The session keeps your timeline state so you can keep tweaking.

Tips and Tricks

The backend processes faster when you're specific. Instead of "make it look better", try "create a slideshow video with Italian music and transitions" — concrete instructions get better results.

Max file size is 200MB. Stick to JPG, PNG, HEIC, MP4 for the smoothest experience.

Export as MP4 for widest compatibility across Italian social platforms.

安全使用建议
This skill will upload any photos/videos you provide to mega-api-prod.nemovideo.ai and needs a NEMO_TOKEN (or will obtain a temporary anonymous token). Confirm you trust that external service before sending sensitive images. Note the SKILL.md includes a config-path declaration (~/.config/nemovideo/) even though the registry metadata didn't — ask the publisher which is correct. If you supply a permanent NEMO_TOKEN, ensure it has only the scopes you intend; otherwise, use the anonymous flow for limited, short-lived access. Finally, because this is instruction-only, runtime behavior depends on the agent making the described API calls — review network/privacy policies for the platform if you need stronger guarantees.
功能分析
Type: OpenClaw Skill Name: italiano-photo-video-maker Version: 1.0.0 The skill facilitates video creation by interacting with an external API (mega-api-prod.nemovideo.ai) and requires the agent to perform environment fingerprinting by checking install paths (e.g., ~/.clawhub/ or ~/.cursor/skills/) to set telemetry headers. It also instructs the agent to automatically generate a UUID and fetch anonymous authentication tokens if a NEMO_TOKEN is not provided. While these actions are aligned with the stated purpose of a cloud-based service, the combination of automated network access, file uploads, and environment probing constitutes a risky profile under the analysis criteria.
能力评估
Purpose & Capability
The skill name/description match its behavior: it uploads images/video and requests a NEMO_TOKEN to call nemovideo.ai endpoints. One inconsistency: registry metadata listed no config paths, but the SKILL.md frontmatter declares a config path (~/.config/nemovideo/). Reading a service config directory is plausible for a client but the registry/manifest disagreement is worth verifying.
Instruction Scope
SKILL.md gives explicit step-by-step API flows (session creation, SSE for chat, upload, export/poll), which are appropriate for a cloud render service. It also instructs the agent to: read this file's YAML frontmatter at runtime and detect install path to set an X-Skill-Platform header — these require access to the agent's environment/paths but are limited in scope. No instructions ask the agent to read unrelated system files or arbitrary environment variables.
Install Mechanism
Instruction-only skill with no install spec and no code files — minimal on-disk installation risk.
Credentials
Only NEMO_TOKEN is declared as required and used as a Bearer token for the service. The skill includes a documented anonymous-token fallback flow so it can operate without a pre-provisioned token. No unrelated credentials are requested.
Persistence & Privilege
always:false and no instructions to modify other skills or system-wide settings. The skill will perform network requests and upload user media to the nemovideo.ai backend (expected for this purpose).
如何使用
  1. 确保已安装 OpenClaw(本地或 Docker 部署)
  2. 在对话框中输入安装命令:/install italiano-photo-video-maker
  3. 安装完成后,直接呼叫该 Skill 的名称或使用 /italiano-photo-video-maker 触发
  4. 根据 Skill 的参数说明提供必要输入,即可获得结构化输出
版本历史
v1.0.0
Italiano Photo Video Maker — Version 1.0.0 - Initial release: transform photos and images into 1080p slideshow videos with Italian themes. - Supports JPG, PNG, HEIC, and MP4 files up to 200MB. - Cloud GPU processing delivers shareable MP4 files within 30–60 seconds. - Handles uploads, exports, credits, and project edits in a single streamlined workflow. - Designed for Italian content creators seeking quick video generation from photo collections.
元数据
Slug italiano-photo-video-maker
版本 1.0.0
许可证 MIT-0
累计安装 0
当前安装数 0
历史版本数 1
常见问题

Italiano Photo Video Maker 是什么?

create photos and images into photo slideshow video with this italiano-photo-video-maker skill. Works with JPG, PNG, HEIC, MP4 files up to 200MB. Italian con... 它是一个面向 Claude Code / OpenClaw 的 AI Agent Skill 插件,目前累计下载 71 次。

如何安装 Italiano Photo Video Maker?

在 OpenClaw 或 Claude Code 对话框中运行命令「/install italiano-photo-video-maker」即可一键安装,无需额外配置。

Italiano Photo Video Maker 是免费的吗?

是的,Italiano Photo Video Maker 完全免费,采用 MIT-0 许可证,可自由下载、安装和使用。

Italiano Photo Video Maker 支持哪些平台?

Italiano Photo Video Maker 跨平台运行,可在任意部署了 OpenClaw / Claude Code 的环境中使用(cross-platform)。

谁开发了 Italiano Photo Video Maker?

由 whitejohnk-26(@whitejohnk-26)开发并维护,当前版本 v1.0.0。

💬 留言讨论