gpt-2

Tag

Cards List
#gpt-2

Transformer Explainer: Interactive Learning of Text-Generative Models

Papers with Code Trending · 2024-08-08 Cached

Transformer Explainer is an interactive visualization tool that allows non-experts to understand the inner workings of the GPT-2 model through real-time experimentation and visualization in a web browser.

0 favorites 0 likes
#gpt-2

Image GPT

OpenAI Blog · 2020-06-17 Cached

OpenAI's Image GPT (iGPT) applies GPT-2 transformers to pixel sequences for image generation and classification, demonstrating that the same architecture used for language can learn coherent visual features in an unsupervised manner and achieve competitive performance on image classification benchmarks.

0 favorites 0 likes
#gpt-2

GPT-2: 1.5B release

OpenAI Blog · 2019-11-05 Cached

OpenAI releases GPT-2 1.5B model with analysis of human perception of credibility, potential for misuse through fine-tuning on extremist ideologies, and challenges in detecting synthetic text. Detection models achieve ~95% accuracy but require complementary approaches for practical deployment.

0 favorites 0 likes
#gpt-2

Fine-tuning GPT-2 from human preferences

OpenAI Blog · 2019-09-19 Cached

OpenAI demonstrates fine-tuning GPT-2 (774M parameters) using human preference feedback for text continuation and summarization tasks, requiring 5k labels for stylistic tasks and 60k for summarization, with models achieving 86-88% human preference rates though revealing labeler heuristic exploitation.

0 favorites 0 likes
#gpt-2

GPT-2: 6-month follow-up

OpenAI Blog · 2019-08-20 Cached

OpenAI discusses their 6-month follow-up to GPT-2 release, outlining plans to release the 1558M parameter model in a few months and emphasizing staged release and partnership-based sharing as key to responsible AI publication.

0 favorites 0 likes
#gpt-2

Better language models and their implications

OpenAI Blog · 2019-02-14 Cached

OpenAI introduces GPT-2, a 1.5 billion parameter transformer-based language model trained on 40GB of internet text that achieves state-of-the-art performance on language modeling benchmarks and demonstrates zero-shot capabilities in reading comprehension, translation, question answering, and summarization. Due to safety concerns, only a smaller model and technical paper are released publicly rather than the full trained model.

0 favorites 0 likes
#gpt-2

togatoga/karukan

GitHub Trending (daily) · 2026-07-01 Cached

Karukan is a neural kana-kanji conversion input method system for Linux/macOS, using llama.cpp to run the GPT-2 model, supporting real-time conversion, context awareness, and user learning.

0 favorites 0 likes
← Previous
← Back to home

Submit Feedback