Skip to content
View Ephemeral182's full-sized avatar
🤪
🤪

Block or report Ephemeral182

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Ephemeral182/README.md
Typing SVG
Profile Views

🧑‍🎓 About Me

🎓 Second-year PhD Student at HKUST (Guangzhou), Robotics and Autonomous Systems (ROAS)
🔬 Research Interests: Generative Models, Reward Modeling, Vision Agents, Visual Perception
🧠 Current Focus: High-quality post-training, human-aligned reward modeling, multimodal agentic RL, and unified reasoning for visual understanding and generation
🌐 Website: ephemeral182.github.io
📧 Email: ephemeral182@gmail.com

I’m Sixiang Chen (陈思翔), advised by Prof. Lei Zhu and Prof. Fugee Tsung.
My work explores how generative and unified models can unlock new potential in image understanding and creation.

I believe research should matter to real users.
Some projects I lead — especially the PosterX series such as PosterCraft, PosterOmni, and PosterReward — aim to build dependable tools for designers and broader creative communities, while also pushing forward general post-training and alignment for image-text foundation models.


🚀 Research Themes

  • 🧠 High-Quality Post-Training for Generative Models
    Building post-training and alignment methods for high-quality image-text generation and editing, with a focus on reward modeling, preference learning.

  • 📏 Reward Modeling and Preference Alignment
    Designing reward models that align AI outputs with human standards in aesthetics, structure, typography, and usability.

  • 🤖 Vision Agents & Unified Reasoning
    Exploring multimodal agentic RL and unified models for open-world visual understanding, generation, and reasoning.

  • 🧩 Agent Systems & Harness Engineering Studying tool-using coding agents, skill/plugin ecosystems, memory, evaluation, and reliable orchestration through reproducible implementations.


🛠️ Featured Projects

🎯 PosterOmni · CVPR 2026

A unified framework for generalized multi-task poster creation and editing, covering both local refinement and global design.

🧠 PosterReward · CVPR 2026

A reward model for accurate evaluation and alignment of high-quality graphic design generation.

🎨 PosterCraft · ICLR 2026

Rethinking high-quality aesthetic poster generation in a unified framework.

Self-evolving image generation agents through tool-orchestrated visual experience distillation.

An evidence-aware daily Agent research radar with Chinese paper briefings powered by DeepSeek.


🧭 Repository Map

Track Repositories Focus
Agent research GenEvolve · GenEvolve-Training · Frontier Signal Self-evolving agents, multimodal RL, paper intelligence
Agent harness lab DeepSeek Harness · Pi · OpenHarness · Learn Claude Code Agent loops, tools, plugins, orchestration
Skills & plugins ASu Resume Skills · Pi Skills · Awesome Harness Engineering Reusable skills, plugin patterns, learning references
Generative vision PosterCraft · MeSa-IR · SnowFormer Poster generation, image restoration, evaluation

Forks in the Agent harness lab are curated as a study shelf. Original research and maintained projects remain the primary portfolio.


🧪 Agent Learning Path

  1. Core loop — read Pi and Learn Claude Code for compact agent-loop implementations.
  2. Plugin architecture — compare DeepSeek Harness, OpenHarness, and Oh My Pi.
  3. Skills ecosystem — inspect Pi Skills, Awesome Pi Agent, and Awesome DeepSeek Agent.
  4. Harness engineering — use Awesome Harness Engineering as the index for memory, evaluation, permissions, observability, and orchestration.

🌟 Selected Highlights

  • Feb 2026PosterOmni accepted by CVPR 2026
  • Feb 2026PosterReward accepted by CVPR 2026
  • Jan 2026PosterCraft accepted by ICLR 2026
  • Jun 2025GenHaze accepted by ICCV 2025
  • Apr 2025 — Released the GPT-4o Image Generation Capabilities evaluation report
  • 2025 — Multiple papers accepted by CVPR / ICCV / AAAI
  • 2024 — Papers accepted by NeurIPS / ECCV / CVPR, including CVPR 2024 Highlight

📊 GitHub Summary

Ephemeral182's github stats Ephemeral182's Top Languages

🔥 Current Keywords

High-Quality Post-Training · Reward Modeling · Vision Agents · Agent Harness · Skills & Plugins · Image-Text Generation


🧰 Research / Workflow Stack

Type Tools
Writing / Academic Overleaf Notion Excalidraw
AI Research Assistants ChatGPT Claude Perplexity Google Scholar
Development / Coding Cursor Claude Code GitHub Copilot Jina
Training / Infra PyTorch Diffusers Transformers Slurm
Design / Prototyping Stitch draw.io Clipdrop Shields.io Figma

📫 Let's Connect!

🌐 Website: ephemeral182.github.io
🐙 GitHub: Ephemeral182
📧 Email: ephemeral182@gmail.com


WeChat GitHub followers Email

Pinned Loading

  1. MeiGen-AI/PosterOmni MeiGen-AI/PosterOmni Public

    [CVPR2026] PosterOmni: One model for poster creation—unifying local edits and global design for generalized multi-task image/poster-to-poster generation.

    Python 212 12

  2. MeiGen-AI/PosterCraft MeiGen-AI/PosterCraft Public

    [ICLR2026] Rethinking High-Quality Aesthetic Poster Generation in a Unified Framework

    Python 1k 61

  3. PosterCraft PosterCraft Public

    [ICLR'26] Rethinking High-Quality Aesthetic Poster Generation in a Unified Framework

    Python 545 35

  4. MeiGen-AI/GenEvolve MeiGen-AI/GenEvolve Public

    Self-Evolving Image Generation Agents via Tool-Orchestrated Visual Experience Distillation

    Python 86 2

  5. UDR-S2Former_deraining UDR-S2Former_deraining Public

    [ICCV'23] Sparse Sampling Transformer with Uncertainty-Driven Ranking for Unified Removal of Raindrops and Rain Streaks

    Python 149 7

  6. Empirical-Study-of-GPT-4o-Image-Gen Empirical-Study-of-GPT-4o-Image-Gen Public

    An Empirical Study of GPT-4o Image Generation Capabilities

    29