Any-to-Any
MLX
Safetensors
gemma4
mlx-vlm
rlcd
multimodal
classification
parallel-inference
image-text-to-text
audio
video
4-bit precision
Instructions to use larkooo/gemma-e2b-rlcd with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- MLX
How to use larkooo/gemma-e2b-rlcd with MLX:
# Download the model from the Hub pip install huggingface_hub[hf_xet] huggingface-cli download --local-dir gemma-e2b-rlcd larkooo/gemma-e2b-rlcd
- Notebooks
- Google Colab
- Kaggle
- Local Apps Settings
- LM Studio
- Atomic Chat
File size: 1,870 Bytes
53e24ca | 1 2 3 4 5 6 7 8 9 10 11 12 13 14 15 16 17 18 19 20 21 22 23 24 25 26 27 28 29 30 31 32 33 34 35 36 37 38 39 40 41 42 43 44 45 46 47 48 49 50 51 52 53 54 55 56 57 58 59 60 61 62 63 64 65 66 67 68 69 70 71 72 73 74 75 76 77 78 79 80 81 82 83 | """Create synthetic integration fixtures without using personal media."""
import argparse
import subprocess
from pathlib import Path
from PIL import Image
def main():
parser = argparse.ArgumentParser()
parser.add_argument("--output-dir", required=True, type=Path)
args = parser.parse_args()
root = args.output_dir.resolve()
root.mkdir(parents=True, exist_ok=True)
Image.new("RGB", (224, 224), (255, 0, 0)).save(root / "red.png")
subprocess.run(
[
"say",
"-v",
"Samantha",
"-o",
str(root / "dog.wav"),
"--file-format=WAVE",
"--data-format=LEI16@16000",
"A dog is sleeping on the sofa.",
],
check=True,
)
subprocess.run(
[
"ffmpeg",
"-v",
"error",
"-y",
"-f",
"lavfi",
"-i",
"color=c=red:s=224x224:r=2:d=2",
"-f",
"lavfi",
"-i",
"color=c=blue:s=224x224:r=2:d=2",
"-filter_complex",
"[0:v][1:v]concat=n=2:v=1:a=0[v]",
"-map",
"[v]",
"-c:v",
"libx264",
"-pix_fmt",
"yuv420p",
str(root / "red-blue.mp4"),
],
check=True,
)
subprocess.run(
[
"ffmpeg",
"-v",
"error",
"-y",
"-i",
str(root / "red-blue.mp4"),
"-i",
str(root / "dog.wav"),
"-c:v",
"copy",
"-c:a",
"aac",
"-af",
"apad",
"-t",
"4",
str(root / "red-blue-speech.mp4"),
],
check=True,
)
print(root)
if __name__ == "__main__":
main()
|