Yue Zhou

66 posts

Yue Zhou

Yue Zhou

@DataFleuret

PhD Student in NLP@UIC. #NLProc for Social Good, Healthcare #NLP #DeepLearning Amateur Foil Fencing Player. Classical Guitar. Snowboarding.

Chicago, IL Katılım Mart 2019
112 Takip Edilen21 Takipçiler
Yue Zhou retweetledi
Michael Black
Michael Black@Michael_J_Black·
In the LLM-science discussion, I see a common misconception that science is a thing you do and that writing about it is separate and can be automated. I’ve written over 300 scientific papers and can assure you that science writing can’t be separated from science doing. Why? 1/18
English
38
460
1.8K
0
Yue Zhou retweetledi
UIC CS Department
UIC CS Department@UICCS·
Congratulations to Barbara Di Eugenio for receiving the @AWISNational Zenith Award for her lifetime achievements in STEM and her commitment to workplace diversity. bit.ly/3gMgU6G
UIC CS Department tweet media
English
0
3
18
0
Yue Zhou retweetledi
Tal Linzen
Tal Linzen@tallinzen·
Just to make sure this point doesn't get lost: the success of instruction-finetuning highlights the *limitation* of the self-supervised language modeling objective.
Quoc Le@quocleix

New open-source language model from Google AI: Flan-T5 🍮 Flan-T5 is instruction-finetuned on 1,800+ language tasks, leading to dramatically improved prompting and multi-step reasoning abilities. Public models: bit.ly/3sbNPDJ Paper: arxiv.org/abs/2210.11416

English
2
24
207
0
Yue Zhou retweetledi
Stable Diffusion
Stable Diffusion@StableDiffusion·
Over the 7 weeks since Stable Diffusion's release, we've seen many amazing open-source contributions from the community. A lot of them have come in the form of awesome Google Colab notebooks! 🔥 Here is a thread of 14 awesome notebooks we've seen from the community ↓
English
53
662
3.4K
0
Yue Zhou retweetledi
Nanyun (Violet) Peng @ ACL26
Nanyun (Violet) Peng @ ACL26@VioletNPeng·
Excited to share our #Neurips2022 paper on controllable text generation with theoretical guarantees. We decompose a sequence-level oracle into token-level guidance to steer the generation to consider future constraints. Impressive results for incorporating OOVs.
Nanyun (Violet) Peng @ ACL26 tweet media
Tao Meng@TaoMeng10

Happy to share our work on Controllable Text Generation with NeurAlly-Decomposed Oracle (NADO) accepted at #Neurips2022! Looks like our model is excited about NeurIPS as much as we do. See more in arxiv.org/abs/2205.14219

English
2
16
92
0
Yue Zhou retweetledi
Hanna Hajishirzi
Hanna Hajishirzi@HannaHajishirzi·
My alma mater, Sharif University of Technology, Iran's premier university, was under siege yesterday. Many students were arrested, heavily injured or perhaps killed. Many Iranian scholars and PhD students whom you know have received their BSc degrees from this university.
English
1
50
364
0
Yue Zhou
Yue Zhou@DataFleuret·
@chrmanning Even though the generated text doesn’t make much sense…I can still imagine your voice when I am reading it 😂 especially from “of course you can! This is like asking..” would be even better if it ended with “whatsoever” 😀
English
0
0
0
0
Yue Zhou retweetledi
Samuel "curry-howard fanboi" Ainsworth
📜🚨📜🚨 NN loss landscapes are full of permutation symmetries, ie. swap any 2 units in a hidden layer. What does this mean for SGD? Is this practically useful? For the past 5 yrs these Qs have fascinated me. Today, I am ready to announce "Git Re-Basin"! arxiv.org/abs/2209.04836
GIF
English
61
560
2.7K
0
Yue Zhou retweetledi
Chris Olah
Chris Olah@ch402·
I've never had so many "this can't possibly be true, we must have a bug" results in the course of a research project before. I'd like to take a moment to walk through some of the very strange (and surprisingly beautiful) things we found.
Anthropic@AnthropicAI

Neural networks often pack many unrelated concepts into a single neuron – a puzzling phenomenon known as 'polysemanticity' which makes interpretability much more challenging. In our latest work, we build toy models where the origins of polysemanticity can be fully understood.

English
16
210
1.6K
0
Yue Zhou retweetledi
Graham Neubig
Graham Neubig@gneubig·
We've started the Fall 2022 edition of: 🎓CMU CS11-711 Advanced NLP!🎓 Follow along for * An intro of core topics * Timely content; prompting, retrieval, bias/fairness * Content on NLP research methodology Page: phontron.com/class/anlp2022/ Videos: youtube.com/playlist?list=…
Graham Neubig tweet media
English
9
191
755
0
Yue Zhou retweetledi
Sam Bowman
Sam Bowman@sleepinyourhat·
I've had a similar experience to Ethan. If you want to do an NLP data collection/labeling process and don't want/need to be managing annotators directly, Surge is remarkably easy to work with and their team does very good work.
Ethan Perez@EthanJPerez

The biggest game-changer for my research recently has been using @HelloSurgeAI for human data collection. With Surge, the workflow for collecting human data now looks closer to “launching a job on a cluster” which is wild to me. 🧵 of examples:

English
1
7
54
0
Yue Zhou retweetledi
David McClure
David McClure@clured·
Inspired by various recent efforts to make sense of the text2img datasets - here's all 12M captions from LAION-Aesthetics with score > 6, embedded with CLIP and UMAP'ed to 2d. Color is the domain of the image URL.
David McClure tweet media
English
34
274
1.6K
0
Yue Zhou retweetledi
Richard Socher
Richard Socher@RichardSocher·
Nobody is working on this seriously. Large language models have no goals, they just learn to predict the next item in a sequence. We, as a species and research community, are surprisingly uncreative when it comes to identifying useful objective function spaces.
English
3
2
19
0
Yue Zhou retweetledi
Mark Tenenholtz
Mark Tenenholtz@marktenenholtz·
DALLE-2 was paywall-released recently by an extremely well-funded company. Just yesterday, a group of independent researchers released their own model (Stable Diffusion) that you can use in a few lines of code for FREE. The speed of the ML research community is insane 🤯
Mark Tenenholtz tweet media
English
65
997
7.8K
0