zhangyiss/mlx - mlx - Gitea for Geophysics

mirror of https://github.com/ml-explore/mlx.git synced 2025-06-28 12:13:21 +08:00

Author	SHA1	Message	Date
Angelos Katharopoulos	0359bf02c9	Nearest upsample (#2202 )	2025-05-19 11:23:38 -07:00
Awni Hannun	aa5d84f102	Allow quant layer to be unfrozen (#2142 )	2025-04-30 09:08:29 -07:00
Param Thakkar	600e87e03c	Added output_padding parameters in conv_transpose (#2092 )	2025-04-23 09:26:33 -07:00
Angelos Katharopoulos	4eef8102c9	Distributed layers (#1270 )	2025-03-21 13:52:17 -07:00
jiyzhang	65a38c452b	update the formula of smooth_l1_loss (#1986 )	2025-03-21 06:25:23 -07:00
jiyzhang	95e335db7b	Update smooth_l1_loss in losses.py (#1974 ) According the definition of smooth_l1_loss, the line diff = predictions - targets Should be updated to diff = mx.abs(predictions - targets) After the modification, the result is consistent with PyTorch smooth_l1_loss	2025-03-19 20:19:02 -07:00
Chunyang Wen	cffceda6ee	Add type hint for _extra_repr (#1948 )	2025-03-10 06:05:36 -07:00
Chunyang Wen	d14c9fe7ea	Add file info when raising errors in save (#1943 )	2025-03-08 14:51:04 -08:00
Chunyang Wen	5db90ce822	Fix obsured warning (#1944 )	2025-03-08 14:50:39 -08:00
Chunyang Wen	d699cc1330	Fix unreachable warning (#1939 ) * Fix unreachable warning * Update error message	2025-03-07 17:23:04 -08:00
Angelos Katharopoulos	5d68082881	Ring docs (#1829 )	2025-02-28 11:34:21 -08:00
Franck Verrot	7df3f792a2	Ensure Conv2D and Conv3D's kernel sizes aren't trimmed (#1852 ) Before the change, this snippet: ``` print(nn.Conv1d(1, 32, 3, padding=1)) print(nn.Conv2d(1, 32, (3, 3), padding=1)) print(nn.Conv3d(1, 32, (3, 3, 3), padding=1)) ``` would output: ``` Conv1d(1, 32, kernel_size=3, stride=1, padding=1, dilation=1, groups=1, bias=True) Conv2d(1, 32, kernel_size=(3,), stride=(1, 1), padding=(1, 1), dilation=1, groups=1, bias=True) Conv3d(1, 32, kernel_size=(3, 3), stride=(1, 1, 1), padding=(1, 1, 1), dilation=1, bias=True) ``` After the change, the output will be: ``` Conv1d(1, 32, kernel_size=3, stride=1, padding=1, dilation=1, groups=1, bias=True) Conv2d(1, 32, kernel_size=(3, 3), stride=(1, 1), padding=(1, 1), dilation=1, groups=1, bias=True) Conv3d(1, 32, kernel_size=(3, 3, 3), stride=(1, 1, 1), padding=(1, 1, 1), dilation=1, bias=True) ```	2025-02-10 06:27:01 -08:00
Awni Hannun	ca305afdbe	loading empty list is ok when strict = false (#1834 )	2025-02-05 16:19:27 -08:00
Awni Hannun	1017ac4a9e	add dilation for conv 3d layers + test for 3d conv w/ dilation (#1802 )	2025-01-28 06:17:07 -08:00
Awni Hannun	2235dee906	catch stream errors earlier to avoid aborts (#1801 )	2025-01-27 14:05:43 -08:00
Nripesh Niketan	5cc5201914	feat: Add orthogonal initializer and corresponding tests (#1651 ) * feat: Add orthogonal initializer and corresponding tests * lint * Add acknowledgements * nits --------- Co-authored-by: Awni Hannun <awni@apple.com>	2025-01-13 07:29:20 -08:00
Awni Hannun	657f466402	use sdpa and exportable functions in transformer multi head attention (#1760 )	2025-01-09 13:11:55 -08:00
Awni Hannun	c3628eea49	Add `mx.finfo` and use it when making causal mask (#1726 ) * finfo * fixes * docs	2024-12-19 14:52:41 -08:00
Tomohiro Oga	a6b426422e	add cubic to type hinting for upsample (#1709 )	2024-12-17 07:30:23 -08:00
Awni Hannun	29a620cab2	No reshapes in quantized embedding (#1682 ) * no reshapes in quantized embedding * fix inadvertant cast * add tol	2024-12-09 18:57:38 -08:00
Alex Barron	1445dcaa60	let class predicate specify quantization parameters (#1638 )	2024-12-02 14:09:28 -08:00
Awni Hannun	aa86876813	fix transformer decoder post norm LN (#1637 )	2024-12-02 07:02:17 -08:00
Awni Hannun	7cbb4aef17	Doc fix (#1615 )	2024-11-22 11:12:25 -08:00
Angelos Katharopoulos	d8c824c594	Formatting fixes (#1606 )	2024-11-20 15:30:36 -08:00
Saanidhya	cb431dfc9f	Adds 3D pooling (#1526 )	2024-11-19 16:45:24 -08:00
Awni Hannun	59247c2b62	add groups in conv2d (#1569 )	2024-11-07 13:57:53 -08:00
Awni Hannun	4f72c66911	improvements to scatter / gather (#1541 )	2024-10-30 19:30:54 -07:00
Venkata Naga Aditya Datta Chivukula	430ffef58a	[Feature] Added Sparse Initialization (#1498 ) Co-authored-by: Saanidhyavats <saanidhyavats@gmail.com>	2024-10-24 12:31:24 -07:00
LastWhisper	2b8ace6a03	Typing the dropout. (#1479 )	2024-10-15 06:45:46 -07:00
Lucas Newman	4a64d4bff1	Add support for grouped 1D convolutions to the nn API (#1444 ) * Fix the weight shape for grouped convolutions from the nn API. * Add tests. * Pre-commit formatting. * Add input validation. * Use integer division instead of casting. * docs * nit --------- Co-authored-by: Awni Hannun <awni@apple.com>	2024-09-28 06:41:07 -07:00
Awni Hannun	c6739ba7f3	Faster RNN layers (#1419 ) * faster rnn * use admm	2024-09-17 06:04:19 -07:00
Angelos Katharopoulos	914409fef9	Data parallel helper (#1407 )	2024-09-16 18:17:21 -07:00
c0g	bd8396fad8	Fix typo in transformer docs (#1414 )	2024-09-14 06:05:15 -07:00
Awni Hannun	8b30acd7eb	fix module attribute set, reset, set (#1403 )	2024-09-11 16:30:42 -07:00
Max-Heinrich Laves	efeb9c0f02	Transposed Convolution (#1245 ) * initial implementation for conv_transpose ran pre-commit implemented conv_transpose updated conv_general docstring updated conv_general docstring updated code comments removed commented run_conv_checks updated acknowledgments added missing entry to ops.rst added op to nn.layers resolved merge conflicts * removed ConvolutionTranspose primitive as suggested by reviewer removed ConvolutionTranspose primitive as suggested by reviewer * remove transpose flag, add another test --------- Co-authored-by: Awni Hannun <awni@apple.com>	2024-09-06 19:52:38 -07:00
Awni Hannun	291cf40aca	Some fixes to typing (#1371 ) * some fixes to typing * fix module reference * comment	2024-08-28 11:16:19 -07:00
Awni Hannun	63ae767232	fix transformer (#1327 )	2024-08-13 16:04:26 -07:00
Bhargav Yagnik	a098bc92e0	Fix: Preserve input dtype in Dropout layer output (#1323 ) * Fix: Preserve input dtype in Dropout layer output - Modified Dropout implementation to ensure that the output dtype matches the input dtype. - This resolves the issue #1321 * Update test cases in test_nn.py - Revised test cases to align with updated dropout code - Fixed assertion method: replaced self.assertTrue with self.assertEqual for accurate comparisons in test_nn.py -> test_rope, test_alibi and test_dropout, * updated dropout.py	2024-08-13 11:54:21 -07:00
Alex Barron	635ccd9e25	Add "edge" mode to mx.pad (#1309 ) * Add edge padding mode * fix pad in pooling * string arg instead of enum	2024-08-06 11:23:10 -07:00
Awni Hannun	6c8dd307eb	faster group norm (#1304 )	2024-08-01 12:49:23 -07:00
Atakan Tekparmak	6e06e3a904	feat: Added "tanh" option to GELU approximation (#1268 )	2024-07-28 09:07:56 +02:00
Paul Paczuski	ebd7135b50	Improve stability of BCE loss calculation for input probabilities close to or exactly 0 or 1 (#1280 ) * Improve stability of BCE loss calculation * Standardize comment * Apply formatting with black via pre-commit * Add usage recommendation to docstring * Update python/mlx/nn/losses.py --------- Co-authored-by: Awni Hannun <awni.hannun@gmail.com>	2024-07-24 08:38:22 -07:00
Awni Hannun	20bb301195	CPU binary reduction + Nits (#1242 ) * very minor nits * reduce binary * fix test	2024-06-28 13:50:42 -07:00
Nikhil Mehta	0b7d71fd2f	Add softmin, hardshrink, hardtanh (#1180 ) --------- Co-authored-by: Nikhil Mehta <nikmehta@tesla.com>	2024-06-04 15:48:18 -07:00
Dominik Schlösser	3576b547c5	Doc error for default for scale in SinusoidalPositionalEncoding (#1174 )	2024-06-02 13:42:45 -07:00
Awni Hannun	e6fecbb3e1	Some fixes in docs (#1141 ) * fixes in docs * nit	2024-05-20 11:51:47 -07:00
Angelos Katharopoulos	e78a6518fa	Block sparse qmm (#1124 )	2024-05-16 15:24:14 -07:00
Cheng	5be5daa6ef	Use compiled function in Sigmoid module (#1116 )	2024-05-14 06:25:57 -07:00
Cheng	60cb11764e	Use correct module type in quantized.py (#1115 )	2024-05-14 06:25:42 -07:00
Max-Heinrich Laves	ff4223904d	Conv3d (#993 ) * added conv3d added conv3d implemented explicit_gemm_conv_ND_cpu and bounds checks for slow_conv_3D * incorporated reviewer comments * fixed test * reduced tensor shapes in test for conv3d * Reviewer suggestion Co-authored-by: Awni Hannun <awni.hannun@gmail.com> Reviewer suggestion Co-authored-by: Awni Hannun <awni.hannun@gmail.com> Reviewer suggestion Co-authored-by: Awni Hannun <awni.hannun@gmail.com> Reviewer suggestion	2024-05-11 06:15:02 -07:00

1 2 3 4

155 Commits