Wayback Machine
8 captures
19 Jul 2020 - 07 Jan 2025
Jun JUL Aug
20
2019 2020 2021
success
fail
About this capture
TIMESTAMPS
loading
The Wayback Machine - https://web.archive.org/web/20200720035441/https://twitter.com/sh_reya/status/1284545976892403714
Skip to content
  • Home Home Home, current page.
  • Moments Moments Moments, current page.
  • Language: English
    • Bahasa Indonesia
    • Bahasa Melayu
    • Català
    • Čeština
    • Dansk
    • Deutsch
    • English UK
    • Español
    • Filipino
    • Français
    • Hrvatski
    • Italiano
    • Magyar
    • Nederlands
    • Norsk
    • Polski
    • Português
    • Română
    • Slovenčina
    • Suomi
    • Svenska
    • Tiếng Việt
    • Türkçe
    • Ελληνικά
    • Български език
    • Русский
    • Српски
    • Українська мова
    • עִבְרִית
    • العربية
    • فارسی
    • मराठी
    • हिन्दी
    • বাংলা
    • ગુજરાતી
    • தமிழ்
    • ಕನ್ನಡ
    • ภาษาไทย
    • 한국어
    • 日本語
    • 简体中文
    • 繁體中文
  • Have an account? Log in
    Have an account?

    New to Twitter?
    Sign up
sh_reya's profile
Shreya Shankar
Shreya Shankar
Shreya Shankar
@sh_reya

Tweets

Shreya Shankar

@sh_reya

Trying to make machine learning work in the real world. Currently 1st ML engineer @viaduct_ai. Previously did ML research @googlebrain & CS @Stanford. She/her.

San Francisco, CA
shreya-shankar.com
Joined January 2014

Tweets

  • © 2020 Twitter
  • About
  • Help Center
  • Terms
  • Privacy policy
  • Cookies
  • Ads info
Dismiss
Previous
Next

Go to a person's profile

Promote this Tweet

Block

  • Tweet with a location

    You can add location information to your Tweets, such as your city or precise location, from the web and via third-party applications. You always have the option to delete your Tweet location history. Learn more

    Your lists

    Create a new list


    Under 100 characters, optional

    Privacy

    Copy link to Tweet

    Embed this Tweet

    Embed this Video

    Add this Tweet to your website by copying the code below. Learn more

    Add this video to your website by copying the code below. Learn more

    By embedding Twitter content in your website or app, you are agreeing to the Twitter Developer Agreement and Developer Policy.

    Preview

    Why you're seeing this ad

    Log in to Twitter

    Don't have an account? Sign up »

    Sign up for Twitter

    Not on Twitter? Sign up, tune into the things you care about, and get updates as they happen.

    Sign up
    Have an account? Log in »

    Two-way (sending and receiving) short codes:

    Country Code For customers of
    United States 40404 (any)
    Canada 21212 (any)
    United Kingdom 86444 Vodafone, Orange, 3, O2
    Brazil 40404 Nextel, TIM
    Haiti 40404 Digicel, Voila
    Ireland 51210 Vodafone, O2
    India 53000 Bharti Airtel, Videocon, Reliance
    Indonesia 89887 AXIS, 3, Telkomsel, Indosat, XL Axiata
    Italy 4880804 Wind
    3424486444 Vodafone
    » See SMS short codes for other countries

    Confirmation

     

    Welcome home!

    This timeline is where you’ll spend most of your time, getting instant updates about what matters to you.

    Tweets not working for you?

    Hover over the profile pic and click the Following button to unfollow any account.

    Say a lot with a little

    When you see a Tweet you love, tap the heart — it lets the person who wrote it know you shared the love.

    Spread the word

    The fastest way to share someone else’s Tweet with your followers is with a Retweet. Tap the icon to send it instantly.

    Join the conversation

    Add your thoughts about any Tweet with a Reply. Find a topic you’re passionate about, and jump right in.

    Learn the latest

    Get instant insight into what people are talking about now.

    Get more of what you love

    Follow more accounts to get instant updates about topics you care about.

    Find what's happening

    See the latest conversations about any topic instantly.

    Never miss a Moment

    Catch up instantly on the best stories happening as they unfold.

    Shreya Shankar‏ @sh_reya Jul 18

    Got my invite to the @OpenAI GPT-3 API from @gdb. I actually think it deserves more hype than it’s getting, but not necessarily for the magical reasons Twitter touts. Why? My quick thoughts and impressions: (1/11)

    10:50 AM - 18 Jul 2020
    • 555 Retweets
    • 2,805 Likes
    • Suren Aghajanyan Steve Waite Emily Patterson Mihai Serban ✌🏻 David Harris Bogdan Stanciu @rv_inc Shardul Kyle Hardgrave Carson Kahn (🦄 deep neural pony)
    43 replies 555 retweets 2,805 likes
      1. New conversation
      2. Shreya Shankar‏ @sh_reya Jul 18

        First, let me summarize the API documentation. A user has access to 4 models (varying sizes). They cannot fine-tune the models. There’s basically only one function: the user can input some text (“priming”) and the model will predict the next several tokens (words). (2/11)

        2 replies 3 retweets 115 likes
        Show this thread
      3. Shreya Shankar‏ @sh_reya Jul 18

        When I first played with the API, I was stumped. Why couldn’t I replicate stellar Twitter demos? Was everyone just sharing cherry-picking text samples? I wanted the model to identify basic patterns in some unstructured data, but it gave garbage when I input just the data. (3/11)

        1 reply 3 retweets 83 likes
        Show this thread
      4. Shreya Shankar‏ @sh_reya Jul 18

        @notsleepingturk suggested I reformat my inputs as tuples of (unstructured data, example pattern, indicator if the pattern exists). The model could then easily “autocomplete” tuples with missing indicators. Damn. Priming is obviously an art. (4/11)

        5 replies 3 retweets 170 likes
        Show this thread
      5. Shreya Shankar‏ @sh_reya Jul 18

        So why is GPT-3 so hype? It’s amazingly powerful *if* you know how to prime the model well. It’s going to change the ML paradigm — instead of constructing giant train sets for models, we’ll be crafting a few examples for models to do “few-shot” extrapolation from. (5/11)

        1 reply 42 retweets 326 likes
        Show this thread
      6. Shreya Shankar‏ @sh_reya Jul 18

        Shreya Shankar Retweeted Sharif Shameem

        @sharifshameem cracked the skill of priming in his demos. We don’t see what he prepends the demo input with before sending it to the API. Figuring out how to prime models properly will be the key to successfully utilizing language models in the future. (6/11)https://twitter.com/sharifshameem/status/1282676454690451457 …

        Shreya Shankar added,

        2:00
        Sharif Shameem @sharifshameem
        This is mind blowing. With GPT-3, I built a layout generator where you just describe any layout you want, and it generates the JSX code for you. W H A T pic.twitter.com/w8JkrZO4lk
        Show this thread
        3 replies 25 retweets 257 likes
        Show this thread
      7. Shreya Shankar‏ @sh_reya Jul 18

        Let’s deep-dive into how, in theory, one can be a ⭐️ primer. The model’s goal is to maximize the log-likelihood of successive tokens given the primed input. In English, this ends up being: find similar patterns in the training set, and output similar successive tokens. (7/11)

        1 reply 6 retweets 140 likes
        Show this thread
      8. Shreya Shankar‏ @sh_reya Jul 18

        GPT-3 is a great example of the “garbage-in-garbage-out” principle. If you prime poorly, you get shitty results. But since the models are probably trained on basically every piece of data on the Internet, chances are if you prime well, you’ll get intelligent outputs. (8/11)

        4 replies 21 retweets 201 likes
        Show this thread
      9. Shreya Shankar‏ @sh_reya Jul 18

        I like to think of these language models as “children with infinite memory.” Children’s skills are not all that refined, but they have basic pattern-matching skills. Coupled with a superpower to memorize the entire world, well, couldn’t they be extremely useful? (9/11)

        3 replies 31 retweets 273 likes
        Show this thread
      10. Shreya Shankar‏ @sh_reya Jul 18

        What else is so hype? The API’s best model is 350 GB. Serving this monstrosity efficiently and cheaply is an entirely new software problem for the industry. If @OpenAI cracks this, they can become the AWS of modeling. (10/11)

        6 replies 18 retweets 325 likes
        Show this thread
      11. Shreya Shankar‏ @sh_reya Jul 18

        TLDR, if this takes off: 1) Expect the next generation of good ML practitioners to be in way more creative. It’s taking me a while to wrap my head around how to prime this model to get cool demos, lol. 2) Startups will move away from training their own in-house models. (11/11)

        18 replies 31 retweets 515 likes
        Show this thread
      12. End of conversation

    Loading seems to be taking a while.

    Twitter may be over capacity or experiencing a momentary hiccup. Try again or visit Twitter Status for more information.

      Promoted Tweet

      false

      • © 2020 Twitter
      • About
      • Help Center
      • Terms
      • Privacy policy
      • Cookies
      • Ads info