新闻与深度文章
Microsoft has released a new open-source library called DeepSpeed (opens in new tab), which, when combined with its ‘ZeRO’ module can train 100 billion parameter models without using the resources traditionally associated with that.
Transformer-based language generation models have enabled better conversational applications. Though they still have their shortcomings, which were recently exposed by a team at MIT, researchers continue improving them to build better, larger, and more robust models.
新闻报道 | The Register
Meet Clippy 9000: Microsoft brags about building Earth’s largest AI language model, refuses to let it out of the lab
There’s a new giant AI language model in town: enter Microsoft’s Turing-NLG system, which apparently contains a whopping 17 billion parameters, making it the largest publicly known model of its class yet.
新闻报道 | VentureBeat
Microsoft trains world’s largest Transformer language model
Microsoft AI & Research today shared what it calls the largest Transformer-based language generation model ever and open-sourced a deep learning library named DeepSpeed to make distributed training of large models easier.
新闻报道 | InfoWorld
Microsoft speeds up PyTorch with DeepSpeed
Microsoft has released DeepSpeed, a new deep learning optimization library for PyTorch, that is designed to reduce memory use and train models with better parallelism on existing hardware.
There’s an old adage that says if you fail to plan, you plan to fail. But when it comes to AI, Dr. Saleema Amershi, a principal researcher in the Adaptive Systems and Interaction group at Microsoft Research, contends that if…
Meredith (Merrie) Ringel Morris, Sr. Principal Researcher & Research Manager, and Steven Drucker, Partner Research Manager, were both inducted into the CHI Academy. The CHI Academy is an honorary group of individuals who have made substantial contributions to the field…
Microsoft has announced that it has integrated an optimized implementation of BERT (Bidirectional Encoder Representations from Transformers) with the open source ONNX Runtime. Developers can take advantage of this implementation for scalable inferencing of BERT at an affordable cost.
新闻报道 | WinBuzzer
Microsoft Open Sources BERT for ONNX Runtime
In December, Microsoft open sourced its ONNX Runtime inference engine. Now, the company says it also open-sourced an optimized version of BERT, a natural language model from Google, for ONNX.