使用来自 pi 相机的 python 解码并显示 H.264 卡住的视频序列答案

【问题标题】：decode and show H.264 chucked video sequence with python from pi camera使用来自 pi 相机的 python 解码并显示 H.264 卡住的视频序列
【发布时间】：2020-05-16 19:31:17
【问题描述】：

我想解码 H.264 视频序列并将它们显示在屏幕上。视频序列来自 pi 相机，我使用以下代码捕获

import io
import picamera

stream = io.BytesIO()
while True:
    with picamera.PiCamera() as camera:
        camera.resolution = (640, 480)
        camera.start_recording(stream, format='h264', quality=23)
        camera.wait_recording(15)
        camera.stop_recording()

有什么方法可以解码“流”数据序列并使用 opencv 或其他 python 库显示它们？

【问题讨论】：

标签： python opencv raspberry-pi h.264 picamera

【解决方案1】：

我找到了使用 ffmpeg-python 的解决方案。
我无法验证 raspberry-pi 中的解决方案，所以我不确定它是否适合您。

假设：

stream 将整个捕获的 h264 流保存在内存缓冲区中。
您不想将流写入文件。

解决方案应用以下内容：

在子进程中执行FFmpeg，sdtin 作为输入pipe，stdout 作为输出pipe。
输入将是视频流（内存缓冲区）。
输出格式是 BGR 像素格式的原始视频帧。
将流内容写入pipe（至stdin）。
读取解码后的视频（逐帧），并显示每一帧（使用cv2.imshow）

代码如下：

import ffmpeg
import numpy as np
import cv2
import io

width, height = 640, 480


# Seek to stream beginning
stream.seek(0)

# Execute FFmpeg in a subprocess with sdtin as input pipe and stdout as output pipe
# The input is going to be the video stream (memory buffer)
# The output format is raw video frames in BGR pixel format.
# https://github.com/kkroening/ffmpeg-python/blob/master/examples/README.md
# https://github.com/kkroening/ffmpeg-python/issues/156
# http://zulko.github.io/blog/2013/09/27/read-and-write-video-frames-in-python-using-ffmpeg/
process = (
    ffmpeg
    .input('pipe:')
    .video
    .output('pipe:', format='rawvideo', pix_fmt='bgr24')
    .run_async(pipe_stdin=True, pipe_stdout=True)
)


# https://stackoverflow.com/questions/20321116/can-i-pipe-a-io-bytesio-stream-to-subprocess-popen-in-python
# https://gist.github.com/waylan/2353749
process.stdin.write(stream.getvalue())  # Write stream content to the pipe
process.stdin.close()  # close stdin (flush and send EOF)


#Read decoded video (frame by frame), and display each frame (using cv2.imshow)
while(True):
    # Read raw video frame from stdout as bytes array.
    in_bytes = process.stdout.read(width * height * 3)

    if not in_bytes:
        break

    # transform the byte read into a numpy array
    in_frame = (
        np
        .frombuffer(in_bytes, np.uint8)
        .reshape([height, width, 3])
    )

    #Display the frame
    cv2.imshow('in_frame', in_frame)

    if cv2.waitKey(100) & 0xFF == ord('q'):
        break

process.wait()
cv2.destroyAllWindows()

注意：我使用 sdtin 和 stdout 作为管道（而不是使用命名管道），因为我希望代码也可以在 Windows 中工作。

为了测试解决方案，我创建了一个示例视频文件，并将其读入内存缓冲区（编码为 H.264）。
我使用内存缓冲区作为上述代码的输入（替换您的stream）。

这里是完整的代码，包括测试代码：

import ffmpeg
import numpy as np
import cv2
import io

in_filename = 'in.avi'

# Build synthetic video, for testing begins:
###############################################
# ffmpeg -y -r 10 -f lavfi -i testsrc=size=160x120:rate=1 -c:v libx264 -t 5 in.mp4
width, height = 160, 120

(
    ffmpeg
    .input('testsrc=size={}x{}:rate=1'.format(width, height), r=10, f='lavfi')
    .output(in_filename, vcodec='libx264', crf=23, t=5)
    .overwrite_output()
    .run()
)
###############################################


# Use ffprobe to get video frames resolution
###############################################
p = ffmpeg.probe(in_filename, select_streams='v');
width = p['streams'][0]['width']
height = p['streams'][0]['height']
n_frames = int(p['streams'][0]['nb_frames'])
###############################################


# Stream the entire video as one large array of bytes
###############################################
# https://github.com/kkroening/ffmpeg-python/blob/master/examples/README.md
in_bytes, _ = (
    ffmpeg
    .input(in_filename)
    .video # Video only (no audio).
    .output('pipe:', format='h264', crf=23)
    .run(capture_stdout=True) # Run asynchronous, and stream to stdout
)
###############################################


# Open In-memory binary streams
stream = io.BytesIO(in_bytes)

# Execute FFmpeg in a subprocess with sdtin as input pipe and stdout as output pipe
# The input is going to be the video stream (memory buffer)
# The ouptut format is raw video frames in BGR pixel format.
# https://github.com/kkroening/ffmpeg-python/blob/master/examples/README.md
# https://github.com/kkroening/ffmpeg-python/issues/156
# http://zulko.github.io/blog/2013/09/27/read-and-write-video-frames-in-python-using-ffmpeg/
process = (
    ffmpeg
    .input('pipe:')
    .video
    .output('pipe:', format='rawvideo', pix_fmt='bgr24')
    .run_async(pipe_stdin=True, pipe_stdout=True)
)


# https://stackoverflow.com/questions/20321116/can-i-pipe-a-io-bytesio-stream-to-subprocess-popen-in-python
# https://gist.github.com/waylan/2353749
process.stdin.write(stream.getvalue())  # Write stream content to the pipe
process.stdin.close()  # close stdin (flush and send EOF)


#Read decoded video (frame by frame), and display each frame (using cv2.imshow)
while(True):
    # Read raw video frame from stdout as bytes array.
    in_bytes = process.stdout.read(width * height * 3)

    if not in_bytes:
        break

    # transform the byte read into a numpy array
    in_frame = (
        np
        .frombuffer(in_bytes, np.uint8)
        .reshape([height, width, 3])
    )

    #Display the frame
    cv2.imshow('in_frame', in_frame)

    if cv2.waitKey(100) & 0xFF == ord('q'):
        break

process.wait()
cv2.destroyAllWindows()

【讨论】：

这是一个非常好的答案，似乎工作得很好。您对通过调整您发布的代码如何正确阅读直播有什么建议吗？我似乎无法使用streamlink 正确实现它的实时视频。

【解决方案2】：

我认为 OpenCV 不知道如何解码 H264，因此您必须依赖其他库将其转换为 RGB 或 BGR。

另一方面，您可以在picamera中使用format='bgr'，让您的生活更轻松：

Accessing the Raspberry Pi Camera with OpenCV and Python

【讨论】：

【解决方案3】：

我不知道你到底想做什么，但没有 FFMPEG 的另一种方法是：

如果您阅读 picam 文档，您会看到视频端口有 splitters，您可以使用 splitter_port=x (1 start_recorder()：
https://picamera.readthedocs.io/en/release-1.13/api_camera.html#picamera.PiCamera.start_recording

基本上，这意味着您可以将录制的流拆分为 2 个子流，您可以将其中的一个编码为 h264 以进行保存或其他任何操作，以及将其编码为 OPENCV 兼容格式的格式。 https://picamera.readthedocs.io/en/release-1.13/recipes2.html?highlight=splitter#capturing-to-an-opencv-object

这一切主要发生在 GPU 中，因此速度非常快（有关更多信息，请参阅 picamera 文档）

如果您需要一个示例，这与他们在此处所做的相同： https://picamera.readthedocs.io/en/release-1.13/recipes2.html?highlight=splitter#recording-at-multiple-resolutions 但随后有一个 opencv 对象和一个 h264 流

【讨论】：

【解决方案4】：

@Rotem 的回答是正确的，但它不适用于大视频块。

要处理更大的视频，我们需要将process.stdin.write 替换为process.communicate。更新以下几行

...
# process.stdin.write(stream.getvalue())  # Write stream content to the pipe
outs, errs = process.communicate(input=stream.getvalue())
# process.stdin.close()  # close stdin (flush and send EOF)
# Read decoded video (frame by frame), and display each frame (using cv2.imshow)

position = 0
ct = time.time()
while(True):
    # Read raw video frame from stdout as bytes array.
    in_bytes = outs[position: position + width * height * 3]
    position += width * height * 3
...

【讨论】：