主题

图像与视觉

图像上传、视觉分析、隐私边界和图像生成基础。

9 项内容 (6 篇文章 · 3 个视频)

从这里开始

浏览完整内容之前,先阅读几篇推荐文章。

此主题下的更多内容

6 分钟阅读
文章

Visual References Without Style Theft

Reference images and artists are how makers learn - and how generative tools get misused. A consent-aware, fair-reference practice for AI image work that refuses living-artist style cloning while still letting you study craft.

初级
6 分钟阅读
文章

What AI Image Tools Get Wrong, and Why Detectors Don't Fix It

AI-generated and AI-edited images have specific, learnable failure patterns worth knowing. What they don't have is a reliable detector that tells you which category any single image falls into - here is why, and what to do instead.

初级
7 分钟阅读
文章

Get Consent Before You AI-Edit or Share Someone's Photo

Uploading a friend's photo to an AI tool to remove a background, swap a smile, or turn it into an illustration feels like a small, personal edit. For the person in the photo, it can be a much bigger decision they never got to make.

初级
7 分钟阅读
文章

AI语音与音频:从声音克隆到播客再到翻译

2026年的AI音频涵盖四类实用功能:声音克隆、旁白、转录和翻译。本文将带你了解真正好用的工具,以及每一类功能的具体用例。

初级
48 分钟
视频

2024年Midjourney终极新手指南

Future Tech Pilot. 如果你已经决定尝试Midjourney,这份指南会从头到尾带你走完整个流程,包括提示词结构、stylize / chaos / weird参数、图像提示、混合、缩放和平移,以及风格调校。虽然界面仍在不断演变,但这里介绍的内容依然适用,其中形成提示词的思路也能顺畅迁移到v7。

初级
48 分钟
视频

你究竟应该使用哪款AI图像生成器?

Matt Wolfe. Matt使用同一组提示词测试Midjourney、DALL·E、Adobe Firefly、Stable Diffusion、Leonardo等工具,并按照文章关注的维度进行评分,包括准确度、写实度、插画、标志、文字和价格。此后模型阵容已有变化,如今Flux也已加入竞争,但这套评估方法仍使它成为了解这一领域的最佳入门视频。

初级
4 分钟
视频

GPT-4o视觉能力现场演示

OpenAI. 视频用四分钟展示了一个人把手写的线性方程举到摄像头前,ChatGPT在不直接给出答案的情况下辅导他解题。这是最清晰、最简短的演示,能让你切实体会“模型真的能看到我展示的内容”;与我们找到的其他操作演示相比,它也更准确地呈现了文章推荐的使用场景——手写笔记、简单数学题和拍摄的文档。

AI新手