An experimental text-to-image generation system that learns a shared latent representation between vision and language by mapping semantic text embeddings into an image autoencoder’s latent space and reconstructing images from these aligned representations.
Related
alphaparkinc/genpark-multimodal-generative-art-prompt-synthesis-engine-skill: Multimodal generative art prompt synthesis & diffusion renderer (SeaArt style)
Multimodal generative art prompt synthesis & diffusion renderer (SeaArt style)
Greninja9257/LabLLM: A native macOS lab for teaching tiny language models to think — build the architecture, train the weights, and watch a small LLM emerge from scratch, locally on Apple Silicon with custom data, tokenizers, checkpoints, and MLX acceleration.
A native macOS lab for teaching tiny language models to think — build the architecture, train the weights, and watch a small LLM emerge from scratch, locally on Apple Silicon with ...
cneuralnetwork/oplogs: Local-first, open-source experiment tracking for machine learning and agents
Local-first, open-source experiment tracking for machine learning and agents