このマニュアルは現在、英語と簡体字中国語で提供しています。 英語版 · 簡体字中国語版
ドキュメント
Skye Desk manual
Everything Skye Desk does and how to use it – from shooting your photos to using the finished LoRA. Each feature shows its status; only features in the current build have steps.
01
はじめに
This chapter gets Skye Desk open on your Mac and explains projects and the workspace.
Download and first open
The first public build is not out yet; join the waitlist to get one email when it is. When it is out:
- Download Skye Desk from the Download page.
- Drag Skye Desk into Applications.
- Open it. If macOS says it cannot check the app, open System Settings › Privacy & Security.
- Scroll to Security and click Open Anyway next to Skye Desk, then confirm.
- From then on Skye Desk opens normally.
Until the build is signed by Apple, macOS asks once. For Apple silicon Macs; system requirements will be confirmed with the first build; Intel Macs are not tested yet.
Workspace and projects
A workspace is a folder on your Mac that holds your projects. A project is one subject – one product, one character, one style – with its images, captions and training runs. Everything stays as ordinary files in that folder.
A first walk-through
- Create a project and give it a trigger word (a made-up word such as prodshoot).
- Drop 15–30 photos on the Dataset page.
- Write a caption for each image, starting with the trigger word.
- Export a training ZIP, or try a run in demo mode.
- Compare the samples in a blind test.
02
Preparing your images
This chapter covers the photos a LoRA learns from: shooting them, choosing them and importing them.
Shoot with your phone
Most people do not need AI pictures to start – 20 to 30 good phone photos of your subject are enough.
How many photos
| Product, Character | 15–30 |
| 衣服 | 20–40 |
| 食べ物 | 15–25 |
| スタイル | 30 or more |
| A real person (only with their consent) | 20–40 |
| Effect, Material, Interior, Packaging | 20–30 |
Vary everything except the subject
- Angles: front, three-quarter, side, back, from above, and a few close-ups.
- Backgrounds: change them 3–5 times. One background in every photo gets learnt as part of the subject.
- Light: daylight by a window, shade, a lamp – not all the same.
- Distance: wide, medium and detail shots.
Phone settings
- Main camera at 1×; not the ultra-wide lens (it bends shapes).
- Portrait mode off, filters off, beauty effects off.
- Tap to focus on the subject; wipe the lens first.
- Keep the highest resolution; HEIC is fine.
Avoid
- Other brands’ logos and any text you do not want learnt.
- Fingers over the subject; your reflection in shiny surfaces.
- Near-identical shots – ten photos from the same spot teach less than three different ones.
Tips by kind
- 製品
- Turn it a little between shots; include the logo side, the back and the base; one close-up of any engraving or label.
- 衣服
- Use a dress form or a hanger, then a few worn shots (face out of frame is fine); include cuffs, collar, buttons and the fabric up close.
- Character / figure
- Front, three-quarter, side, back; the face and any markings close up; keep the same lighting.
- 食べ物
- Shoot within minutes, before it changes; top-down, 45° and side; one shot cut open if the inside matters.
- インテリア
- Corners and straight-on walls at eye height; daylight and lamp light; avoid people walking through.
- スタイル
- Many different subjects in the same style – the subjects should vary, the style should not.
- A real person
- Written consent first. Varied expressions, light and clothes; no other people in frame.
Get them onto your Mac
- Select the photos on your iPhone and share with AirDrop to your Mac.
- They land in Downloads.
- Drag them onto the Dataset page.
A shooting checklist inside the app, which shows what is still missing after import, is 予定.
Use photos you already have
Pick sharp photos of the subject from different occasions; drop blurry ones, heavy filters and screenshots. 15–30 good photos beat 100 similar ones.
Use AI-generated or reference images
You can also build a dataset from pictures you generated, for example from a design sheet. Check each one for drift – a changed logo or pattern is learnt too.
Working with reference images directly, without training, is the Reference Kit: 予定.
Importing
画像を取り込むには? 利用可能
- まずは1つの被写体を、さまざまな角度から撮った画像を15〜30枚用意します。
- Import(取り込み)をクリックするか、画像・フォルダ・ZIP をドラッグします。
- 名前やキャプションで検索して、目的の画像を探せます。
- Dataset overview(データセット概要)を開いて、不足や偏りを確認します。
- 右側のパネルにキャプション(画像を説明するテキスト)を入力します。
Supported: PNG, JPG, WebP, BMP, HEIC and TIFF images, folders, ZIP files, and same-name TXT or JSON captions. iPhone HEIC photos import as they are.
どこにでもドロップ・ペースト 次のベータで提供
グリッド、サイドバーのプロジェクト、Projects(プロジェクト)ページ、Home(ホーム)ページにドロップするか、⌘Vでペーストします。⌥を押しながらドロップするとコピーではなく移動します。コピー前に重複を検出し、読み込みは取り消せます。
Not in the current build yet – no steps until it is.
03
Organising your dataset
This chapter keeps the dataset tidy so the LoRA learns the right things.
Space でプレビュー 利用可能
- Select an image and press Space.
- Use ← と → to move between images.
- Press Space again to close.
キャプション漏れを表示 利用可能
- Look for the marked images in the grid.
- Choose the Missing captions filter to see only those.
- Write or import the captions.
Finder のように使えるデータセット 次のベータで提供
Return で名前変更、⌘⌫ でゴミ箱へ移動(取り消し可能)、Finder と同じカラータグも使えます。Finder での変更はすぐに反映されます。
Not in the current build yet – no steps until it is.
一括名前変更 次のベータで提供
パターンを選び、適用前に新しい名前をすべて確認できます。キャプションも名前が変わり、1回の取り消しで一括で元に戻せます。
Not in the current build yet – no steps until it is.
データセットのヘルスチェック 次のベータで提供
トリガーワードがない画像、トレーニングには小さすぎる画像、縦横比が極端な画像、ほかとほぼ同じ画像をマークし、それぞれに修正方法を示します。ほぼ重複した画像は、Mac 上で動くモデル(同意した場合のみダウンロード)で内容から見つけられます。画像がアップロードされることはありません。
Not in the current build yet – no steps until it is.
Home(ホーム)から全プロジェクトを検索 次のベータで提供
キャプション、ファイル名、プロンプトで全プロジェクトを検索。各プロジェクトには次のステップが表示されます。
Not in the current build yet – no steps until it is.
04
Writing captions
This chapter explains what to write – and what to leave out.
Leave out what you want the LoRA to learn. Write what should stay changeable. Start every caption with your trigger word.
種類別のキャプションのコツ
- キャラクター
- 書くこと: ポーズ、向き(正面・横)、背景、照明、立っている場所。
書かないこと: キャラクター自体の見た目(形、色、顔、模様)。これはトリガーワードに覚えさせます。 - 製品
- 書くこと: 角度、背景、置いている面、照明、周りの小物。
書かないこと: 製品の形、ロゴの位置、柄、素材。 - 衣服
- 書くこと: 見せ方(トルソー、ハンガー、平置き)、角度、背景。
書かないこと: カット、生地、ステッチ、ボタン、色。 - スタイル
- 書くこと: 各画像の被写体(木、家、船など)。
書かないこと: スタイルそのもの(紙の重なり、柔らかな影、白の上の白)。 - エフェクト
- 書くこと: 下にある物(松ぼっくり、鍵など)とその場面。
書かないこと: 効果そのもの(霜の結晶、光、使い込んだ風合い)。 - 素材
- 書くこと: 素材が何の形になっているか、どう照らされているか。
書かないこと: 素材の質感と色。 - インテリア
- 書くこと: 部屋の間取り、家具、カメラ位置、時間帯。
書かないこと: 壁の仕上げ、床、配色、光の質。 - 食べ物
- 書くこと: 皿、テーブル、角度、切ってあるかどうか。
書かないこと: 料理そのもの(生地、フルーツの並べ方、色)。 - パッケージ
- 書くこと: 周りにある物、置いている面、角度、照明。
書かないこと: ボトルや箱の形、ラベルの位置、キャップ、色。
The trigger word
A short made-up word that calls up your subject, such as prodshoot. Spell it the same way in every caption.
トリガーワードの表記チェック 次のベータで提供
トリガーワード(対象を呼び出す単語)の異なる表記を見つけ、1回でそろえます。
Not in the current build yet – no steps until it is.
キャプションをまとめて修正するには? 利用可能
- Caption tools(キャプションツール)› Find and replace(検索と置換)を開きます。
- 検索する語句を入力します。
- 置き換え後の語句を入力します。
- Preview(プレビュー)をクリックし、変更前と変更後の各行を確認します。
- Apply(適用)をクリックします。Undo(取り消し)でまとめて元に戻せます。
タグのバランスを確認 利用可能
- Open Dataset overview.
- Read the share of images for each tag.
- Click a tag to see those images, and fix captions that lean one way.
自分の AI でキャプションを作るには? 利用可能
- すべての画像を書き出し、任意の AI に画像ごとの .txt を作ってもらいます。
- Import(取り込み)› Import Caption Pack(キャプションパックを取り込む)を選び、ファイルを指定します。
- Write to(書き込み先)に正しいプロジェクトが表示されているか確認します。
- 各行を確認します。MATCH は書き込まれ、SKIP は書き込まれません。
- Write captions(キャプションを書き込む)をクリックします。
データセット全体のタグパネル 次のベータで提供
すべてのタグと使用回数を確認し、多数の画像に対して一括で追加・削除・名前変更ができます。取り消しも可能です。
Not in the current build yet – no steps until it is.
AI のキャプションを1件ずつ確認 次のベータで提供
新旧のキャプションを並べて確認し、チェックしたものだけを残せます。先にバックアップが作成されます。
Not in the current build yet – no steps until it is.
05
Training
This chapter turns the dataset into a LoRA – and explains what is real and what is simulated today.
In the current build, training runs in demo mode: no GPU is rented and nothing is charged.
学習用 ZIP を書き出すには? 利用可能
- すべての画像にキャプションを付けます。空のキャプションがあると書き出せません。
- File(ファイル)メニューを開きます。
- Export Training Package(学習パッケージを書き出す)を選ぶか、次のキーを押します: ⌘⇧E.
- 解凍すると、01.png の隣に 01.txt があります。
- kohya(別の学習ツール)で学習しますか? kohya の仕様に合わせ、リピート回数を名前にしたフォルダを1つ追加して、その中にファイルを入れてください。
デモモード
学習を開始するには? 利用可能
- ステップ数(学習を何回繰り返すか)を設定します。
- クラウド GPU(オンラインで借りる高性能なコンピュータ)を選びます。
- 所要時間と費用の見積もりを確認します。
- Start training(学習を開始)をクリックします。このベータでは学習はシミュレーションで、費用はかかりません。
トレーニングのプリセットとチェックリスト 次のベータで提供
LoRA の種類を選ぶと、用意された設定と実行前のチェックリストが使えます。チェックリストはアドバイスであり、スコアではありません。
Not in the current build yet – no steps until it is.
RunPodでワンクリック学習 ベータ期間中に提供
「最安」か「最速」を選ぶだけで、Skye Deskが空いているGPUを選び、レンタル前に見積もりを表示します。
Not in the current build yet – no steps until it is.
支出上限と自動停止 ベータ期間中に提供
実行ごとに上限を設定。トレーニングの終了時、アイドル時、上限到達時に GPU が停止します。
Not in the current build yet – no steps until it is.
止まったところから再開 ベータ期間中に提供
その場合も、新しい GPU で最後に保存したステップからトレーニングを続けます。
Not in the current build yet – no steps until it is.
Connecting Runpod
Real training on Runpod is ベータ期間中に提供. You will need a Runpod account with credit and an API key, pasted on the Connections page.
Simple and advanced modes
Simple mode asks for the kind, base model and budget; advanced mode shows every setting with an explanation. 予定
Using reference images
Reference Kit:トレーニングの代わりにリファレンス画像を使う 予定
デザインシートからリファレンスセットを作ってチェックし、リファレンス画像を受け付ける画像ツールで使えます。両方の方法を Skye Desk に比較させ、提案を受けることもできます。
Not in the current build yet – no steps until it is.
Creator mode(クリエイターモード)
Creator mode(クリエイターモード) 予定
ブロガー、AIインフルエンサー、ショート動画クリエイター向けに、キャラクターとそのシリーズを中心にしたシンプルなワークスペースです。
Not in the current build yet – no steps until it is.
06
Generating and comparing
This chapter makes new pictures with your LoRA and helps you keep the best version.
ComfyUI の結果を取り込むには? 利用可能
- 画像番号を選びます。
- プロンプトとネガティブプロンプトを書きます。
- ワークフローをコピーし、ComfyUI で実行します。
- 各結果に画像番号の名前を付けてから、Import(取り込み)します。
チェックポイントを比較するには? 利用可能
- 学習結果のサンプルからブラインドテストを開きます。
- A/B の各ペアで、良いと思う画像を選びます。
- Stop here and see results(ここで終了して結果を見る)をクリックすると、ランキングが表示されます。
- 押し間違えたときは Undo last vote(直前の投票を取り消す)で戻せます。
実行を比較し、最終的な LoRA を設定 次のベータで提供
Projects(プロジェクト)ページで2つのトレーニング実行を並べて比較し、ベースモデルを確認しながら一方を最終的な LoRA としてマークします。
Not in the current build yet – no steps until it is.
画像に保存されたワークフロー 次のベータで提供
Skye Desk は画像に保存された ComfyUI ワークフローを読み取り、ベースモデル(LoRA を追加する AI モデル)ごとにまとめます。
Not in the current build yet – no steps until it is.
プロンプトをキャプションの下書きに 次のベータで提供
画像に保存されたプロンプトをキャプションに挿入し、そのまま編集できます。
Not in the current build yet – no steps until it is.
プロンプトをComfyUIに送信 ベータ期間中に提供
すべての画像番号を、それぞれのプロンプトを入力した状態でまとめてComfyUIに送信します。
Not in the current build yet – no steps until it is.
すべての結果のずれをチェック 予定
新しい画像を対象と比較し、ずれている画像や余計な文字・ロゴが入った画像をマークします。残すか外すかはあなたが決めます。
Not in the current build yet – no steps until it is.
Workflows(ワークフロー)ページ:画像をドロップしてワークフローを特定 予定
画像、JSONファイル、ZIPをWorkflows(ワークフロー)ページにドロップすると、各ワークフローを検出して名前を付け、手持ちのものと比較します。
Not in the current build yet – no steps until it is.
07
Exporting and using your LoRA
This chapter covers where the finished LoRA lives and how to use it.
Today you train with your own tool (AI Toolkit or OneTrainer) from the training ZIP; the LoRA file (.safetensors) appears in that tool’s output folder.
- Copy the .safetensors file into ComfyUI/models/loras.
- In ComfyUI, add a Load LoRA node after the model loader.
- Choose your file and start with a strength of 0.7–1.0.
- Put the trigger word at the start of your prompt.
Base model licences differ: some do not allow commercial use of what you make. Check the licence of the base model before you sell or publish.
実行を比較し、最終的な LoRA を設定 次のベータで提供
Projects(プロジェクト)ページで2つのトレーニング実行を並べて比較し、ベースモデルを確認しながら一方を最終的な LoRA としてマークします。
Not in the current build yet – no steps until it is.
08
Keeping your files safe
This chapter explains where your files are and how changes can be undone.
- Images and captions are ordinary files in your workspace folder; captions sit next to images as .txt files.
- Bulk changes show a preview first and can be undone as a whole.
- Back up the workspace folder like any other folder (Time Machine works).
09
Privacy and consent
This chapter covers what leaves your Mac and whose pictures you may train on.
- Your images stay on your Mac. Nothing is uploaded unless you start a cloud run.
- Train on a real person only with that person’s written consent, and never on minors.
- Do not train on other people’s copyrighted work or likeness without permission.
10
リファレンス
Shortcuts, words and fixes in one place.
キーボードショートカット
- ⌘1–⌘8 サイドバーのページを上から順に開きます。
- ⌘K ページ、操作、プロジェクトを検索します。
- Space 選択した画像をプレビューします。 ← と → で画像を切り替えます。
- ⌘, Settings(設定)を開きます。 ⌘⇧F キャプションの検索と置換を開きます。
用語集
- LoRA
- AI 画像モデルに、キャラクター、製品、スタイルなど1つの被写体を覚えさせる小さな追加ファイルです。
- キャプション
- 1枚の画像を説明するテキストです。AI は画像とキャプションを組み合わせて学習します。
- タグ
- キャプション内の短いフレーズです。たとえば「front view」や「white studio background」など。
- トリガーワード
- 新しい画像を作るときに被写体を呼び出す造語です。たとえば wren など。
- ベースモデル
- LoRA を追加する元の AI 画像モデルです。LoRA は学習に使ったベースモデルの系統でしか機能しません。
- チェックポイント
- 学習途中で保存した LoRA のコピーです。チェックポイントを比べると、どのステップが一番良いか分かります。
- クラウド GPU
- 学習を実行するために、時間単位でオンラインで借りる高性能なコンピュータです。
- ComfyUI
- AI モデルで画像を作る無料アプリです。Skye Desk がプロンプトを用意し、ComfyUI が画像を作ります。
- AI Toolkit、OneTrainer
- 無料の学習ソフトです。Skye Desk は、これらに必要なファイルを準備します。
Troubleshooting
- macOS says it cannot check the app
- System Settings › Privacy & Security › Open Anyway (see chapter 01).
- Export is blocked
- Every image needs a caption; use the Missing captions filter.
- Imported captions did not match
- Caption files must share the image’s file name (01.png ↔ 01.txt).
- The LoRA learnt the background
- Re-shoot with 3–5 different backgrounds, or describe the background in each caption.
Nothing found. Try another word.