base64 image decoder free download

Step1X-Edit

A SOTA open-source image editing model

Step1X-Edit is a state-of-the-art open-source image editing model/framework that uses a multimodal large language model (LLM) together with a diffusion-based image decoder to let users edit images simply via natural-language instructions plus a reference image. You supply an existing image and a textual command — e.g. “add a ruby pendant on the girl’s neck” or “make the background a sunset over mountains” — and the model interprets the instruction, computes a latent embedding combining the image content and user intent, then decodes a new image implementing the edit. ...

Downloads: 0 This Week

Last Update: 2025-12-29

See Project

x-transformers

A simple but complete full-attention transformer

A simple but complete full-attention transformer with a set of promising experimental features from various papers. Proposes adding learned memory key/values prior to attending. They were able to remove feedforwards altogether and attain a similar performance to the original transformers. I have found that keeping the feedforwards and adding the memory key/values leads to even better performance. Proposes adding learned tokens, akin to CLS tokens, named memory tokens, that is passed through...

Downloads: 6 This Week

Last Update: 2026-02-12

See Project

GLM-OCR

Accurate × Fast × Comprehensive

...The model’s multimodal capabilities allow it to reason across image and text content holistically, capturing structured and unstructured information from pages that include dense tables, seals, code snippets, and varied document graphics. GLM-OCR integrates a comprehensive SDK and inference toolchain that makes it easy for developers to install, invoke, and embed into production pipelines with simple commands or APIs.

Downloads: 22 This Week

Last Update: 7 days ago

See Project

DocArray

The data structure for multimodal data

DocArray is a library for nested, unstructured, multimodal data in transit, including text, image, audio, video, 3D mesh, etc. It allows deep-learning engineers to efficiently process, embed, search, recommend, store, and transfer multimodal data with a Pythonic API. Door to multimodal world: super-expressive data structure for representing complicated/mixed/nested text, image, video, audio, 3D mesh data. The foundation data structure of Jina, CLIP-as-service, DALL·E Flow, DiscoArt etc. ...

Downloads: 0 This Week

Last Update: 2025-03-21

See Project

Bard API

The unofficial python package that returns response of Google Bard

The Python package returns a response of Google Bard through the value of the cookie. This package is designed for application to the Python package ExceptNotifier and Co-Coder. Please note that the bardapi is not a free service, but rather a tool provided to assist developers with testing certain functionalities due to the delayed development and release of Google Bard's API. It has been designed with a lightweight structure that can easily adapt to the emergence of an official API....

Downloads: 2 This Week

Last Update: 2024-02-24

See Project

DALL-E 2 - Pytorch

Implementation of DALL-E 2, OpenAI's updated text-to-image synthesis

...To train CLIP, you can either use x-clip package, or join the LAION discord, where a lot of replication efforts are already underway. Then, you will need to train the decoder, which learns to generate images based on the image embedding coming from the trained CLIP.

Downloads: 15 This Week

Last Update: 2023-10-19

See Project

Karlo

Text-conditional image generation model based on OpenAI's unCLIP

Karlo is a text-conditional image generation model based on OpenAI's unCLIP architecture with the improvement over the standard super-resolution model from 64px to 256px, recovering high-frequency details only in the small number of denoising steps. We train all components from scratch on 115M image-text pairs including COYO-100M, CC3M, and CC12M. In the case of Prior and Decoder, we use ViT-L/14 provided by OpenAI’s CLIP repository.

Downloads: 2 This Week

Last Update: 2023-06-08

See Project

MAE (Masked Autoencoders)

PyTorch implementation of MAE

MAE (Masked Autoencoders) is a self-supervised learning framework for visual representation learning using masked image modeling. It trains a Vision Transformer (ViT) by randomly masking a high percentage of image patches (typically 75%) and reconstructing the missing content from the remaining visible patches. This forces the model to learn semantic structure and global context without supervision. The encoder processes only the visible patches, while a lightweight decoder reconstructs the full image—making pretraining computationally efficient. ...

Downloads: 0 This Week

Last Update: 2025-10-06

See Project

Deep learning time series forecasting

Deep learning PyTorch library for time series forecasting

...Historically, this repository provided open-source benchmarks and codes for flash flood and river flow forecasting. Full transformer (SimpleTransformer in model_dict): The full original transformer with all 8 encoder and decoder blocks. Requires passing the target in at inference.

Downloads: 0 This Week

Last Update: 2022-08-19

See Project

ALAE

Adversarial Latent Autoencoders

...This design allows the model to learn interpretable latent representations that can be manipulated to control generated image attributes.

Downloads: 0 This Week

Last Update: 2026-03-11

See Project

DETR

End-to-end object detection with transformers

...Unlike traditional computer vision techniques, DETR approaches object detection as a direct set prediction problem. It consists of a set-based global loss, which forces unique predictions via bipartite matching, and a Transformer encoder-decoder architecture. Given a fixed small set of learned object queries, DETR reasons about the relations of the objects and the global image context to directly output the final set of predictions in parallel. Due to this parallel nature, DETR is very fast and efficient.

Downloads: 0 This Week

Last Update: 2021-08-04

See Project

Search Results for "base64 image decoder"

Showing 11 open source projects for "base64 image decoder"

Step1X-Edit

x-transformers

GLM-OCR

DocArray

Bard API

DALL-E 2 - Pytorch

Karlo

MAE (Masked Autoencoders)

Deep learning time series forecasting

ALAE

DETR

Search Results for "base64 image decoder"

Showing 11 open source projects for "base64 image decoder"

Step1X-Edit

x-transformers

GLM-OCR

DocArray

Bard API

DALL-E 2 - Pytorch

Karlo

MAE (Masked Autoencoders)

Deep learning time series forecasting

ALAE

DETR

Related Searches

Related Categories